Software Mind Logo

Software Mind

[8SN] Site Reliability Engineer (SRE) – UI/UX

Posted 6 Days Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in Montréal, QC, CAN
Mid level
In-Office or Remote
Hiring Remotely in Montréal, QC, CAN
Mid level
Support deployment and operations of a production UI service on Kubernetes. Monitor service health, troubleshoot incidents using Splunk and logs, perform first-level UI debugging of Web Components, participate in incident response and CI/CD support, and collaborate with engineering to improve reliability within a client-directed backlog.
The summary above was generated by AI
Company Description

We are Software Mind, an awesome team of engineers who are ready to ramp up any top-notch company’s projects! Our aim? To always be one step ahead. Become part of a multicultural company in constant growth with an excellent work environment certified by Great Place To Work!
 

About the Client

Our client is a leading enterprise software company building highly scalable cloud-native platforms used by organizations around the world. Their engineering teams focus on delivering reliable, secure, and high-performing services while embracing modern DevOps, Kubernetes, and cloud technologies.

You will join a team responsible for ensuring the stability, reliability, and operational excellence of a critical UI service running in production.

#LI-DNI

Job Description

About the Role

We are looking for a Site Reliability Engineer (SRE) – UI/UX to support the deployment, operations, and ongoing maintenance of a production UI service running on Kubernetes.

This role focuses on monitoring service health, troubleshooting production issues, investigating incidents, and ensuring reliable service delivery. You will work closely with engineering and client teams to support production operations and complete work based on a client-directed backlog.

While this role supports a UI-based service, it is not a frontend development position. Working knowledge of Web Components is required to perform first-level debugging of UI-related issues, but deep frontend development expertise is not expected.

What You’ll Do

  • Support the deployment, operations, and ongoing maintenance of production services running on Kubernetes.
  • Monitor service health, availability, and performance.
  • Investigate and troubleshoot production incidents using logs, monitoring, and debugging tools.
  • Perform log analysis and incident debugging using Splunk.
  • Identify service issues and collaborate with engineering teams to support timely resolution.
  • Participate in incident response and production support activities.
  • Perform first-level debugging of UI-related issues involving Web Components.
  • Support service reliability and continuous improvement initiatives.
  • Assist with CI/CD pipelines and cloud-native application operations when needed.
  • Work effectively within a client-directed backlog and established priorities.

Qualifications

Required Qualifications

  • 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Support, or a related role.
  • Hands-on experience supporting the deployment, operations, and ongoing maintenance of production services running on Kubernetes.
  • Experience monitoring service health, troubleshooting production issues, and supporting service reliability.
  • Proficiency with Splunk for log analysis and incident debugging.
  • Experience participating in production incident response and root-cause analysis.
  • Working knowledge of Web Components and the ability to perform first-level debugging of UI-related issues.
  • Strong troubleshooting, analytical, and problem-solving skills.
  • Experience collaborating with software engineering and cross-functional teams.
  • Ability to work independently and effectively within a client-directed backlog.
  • Excellent written and spoken English, at least B2 level.

Additional Information

Preferred Qualifications

  • Experience supporting CI/CD pipelines.
  • Familiarity with multi-tenant services.
  • Experience with cloud-native application operations.
  • Experience supporting high-availability enterprise or SaaS platforms.
  • Familiarity with additional monitoring and observability tools.
  • Experience with cloud platforms such as AWS, Azure, or GCP.
  • Familiarity with container and deployment technologies such as Docker and Helm.

What We Offer

  • Competitive salary and laptop
  • Professional development and training opportunities
  • Work with cutting-edge cloud and container technologies
  • Flexible work arrangements and collaborative team environment
  • Impact on organization-wide digital transformation initiatives

Similar Jobs

3 Hours Ago
Easy Apply
Remote or Hybrid
Canada
Easy Apply
Senior level
Senior level
Marketing Tech • Real Estate • Software • PropTech • SEO
Build and own the analytics foundation: maintain dbt models and Snowflake, create ELT pipelines (Python/Airflow), design semantic layers, ensure data quality and reconciliation, enable self-serve analytics and measurement for cross-functional teams, and support AI-ready data infrastructure.
Top Skills: AirflowAirflow DagsApi IntegrationsBigQueryCi/CdDbtDbt Semantic LayerEltGitMixpanelPosthogPythonRedshiftSalesforceSnowflakeSnowflake CortexSQL
3 Hours Ago
Easy Apply
Remote
Canada
Easy Apply
Senior level
Senior level
Big Data • Fintech • Mobile • Payments • Financial Services
Lead technical strategy and roadmap for Affirm's Online Storage platform. Design and build scalable, multi-region datastore solutions, automation control planes, and operational tooling for hundreds of databases. Drive cross-team collaboration, reliability, monitoring, and mentor engineers while delivering high-impact backend and database features.
Top Skills: AWSDistributed SqlDynamoDBKotlinKubernetesMySQLPgbouncerPostgresProxysqlPythonRds ProxyRedisSparkTidbVitess
4 Hours Ago
Remote
Canada
Mid level
Mid level
Cloud • Fintech • Food • Information Technology • Software • Hospitality
Serve as primary customer contact through onboarding and Go-Live, manage multiple implementations, create and deliver onboarding/training plans, perform remote site/network assessments, document installations and deviations, provide post-live support and troubleshooting, and meet activation goals while consulting in French or English.
Top Skills: Point Of Sale (Pos) SoftwareSalesforce CRMToast Pos

What you need to know about the Montreal Tech Scene

With roots dating back to 1642, Montreal is often recognized for its French-inspired architecture and cobblestone streets lined with traditional shops and cafés. But what truly sets the city apart is how it blends its rich tradition with a modern edge, reflected in its evolving skyline and fast-growing tech industry. According to economic promotion agency Montréal International, the city ranks among the top in North America to invest in artificial intelligence, making it le spot idéal for job seekers who want the best of both worlds.

Key Facts About Montreal Tech

  • Number of Tech Workers: 255,000+ (2024, Tourisme Montréal)
  • Major Tech Employers: SAP, Google, Microsoft, Cisco
  • Key Industries: Artificial intelligence, machine learning, cybersecurity, cloud computing, web development
  • Funding Landscape: $1.47 billion in venture capital funding in 2024 (BetaKit)
  • Notable Investors: CIBC Innovation Banking, BDC Capital, Investissement Québec, Fonds de solidarité FTQ
  • Research Centers and Universities: McGill University, Université de Montréal, Concordia University, Mila Quebec, ÉTS Montréal

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account