DistantJob Logo

DistantJob

Sr DevOps Engineer

Reposted One Month Ago
Remote
Hiring Remotely in CAN
Senior level
Remote
Hiring Remotely in CAN
Senior level
Seeking a Sr DevOps Engineer to design and maintain secure, scalable AWS cloud infrastructure, manage CI/CD pipelines, and improve system reliability across advanced data analytics platforms.
The summary above was generated by AI

“Why did the microservice get promoted?  

Because it scaled under pressure.”

We are recruiting a seasoned DevOps Engineer for one of our valued clients, an innovative technology company specializing in AI-powered brand safety and contextual intelligence solutions. This role focuses on building and maintaining the robust infrastructure foundation that supports their advanced data analytics and content measurement platforms.

The ideal candidate will design, secure, and optimize the operational systems that enable scalable, high-performance applications. This is a pivotal role for someone passionate about automation, system ownership, and creating resilient infrastructure solutions from the ground up.

Key Responsibilities
  • Design, implement, and maintain secure, scalable cloud infrastructure for high-throughput, low-latency applications (mainly AWS, with flexibility for multi-cloud environments)
  • Develop and enhance CI/CD pipelines for efficient, reliable, and consistent deployment processes (GitHub Actions and similar tools)
  • Continuously identify and implement infrastructure improvements across the entire product ecosystem
  • Architect and support serverless solutions using AWS Lambda, ECS Fargate, and event-driven architecture components
  • Establish comprehensive monitoring, alerting, and logging frameworks to ensure system reliability and visibility
  • Manage infrastructure-as-code implementations using CloudFormation, Terraform, and related tools
  • Partner with development teams to determine infrastructure requirements, deployment approaches, and operational toolsets
  • Oversee secrets management and identity systems utilizing AWS IAM and similar platforms
  • Maintain compliance with security and privacy standards, including access controls, encryption protocols, and audit mechanisms
  • Resolve production incidents across all service layers with a focus on rapid response and thorough post-incident analysis
  • Create and maintain automated backup, disaster recovery, and failover systems
  • Research and adopt emerging DevOps methodologies and technologies to enhance platform performance and reliability
Required Qualifications
  • 10+ years of software engineering background with 6+ years focused on DevOps, Site Reliability Engineering, or Infrastructure Engineering
  • Demonstrated experience managing production environments with comprehensive infrastructure responsibilities
  • Extensive AWS expertise, including compute, storage, networking, and identity management services
  • Practical experience with serverless technologies (AWS Lambda, Step Functions, EventBridge, API Gateway, ECS Fargate)
  • Advanced skills in Docker and container orchestration platforms (Kubernetes, ECS, GKE)
  • Proficiency with CI/CD platforms such as GitHub Actions, CircleCI, ArgoCD, or Jenkins
  • Strong scripting capabilities in Bash, Python, or Go for automation and tooling development
  • Experience with observability solutions (Datadog, Prometheus, Grafana, ELK stack)
  • Solid understanding of network design, security frameworks, and zero-trust access architectures
  • Knowledge of secrets management systems and infrastructure-level access policy enforcement
  • Exceptional troubleshooting and root cause analysis abilities
  • Strong collaborative and communication skills across diverse technical and business teams
  • Continuous improvement mindset with focus on automation, optimization, and security enhancement

This position offers the opportunity to join an early-stage team with proven success in developing cutting-edge contextual intelligence technology, working alongside talented engineers and product leaders who prioritize innovation and measurable client impact.

What are you waiting for? Fill out the form below and apply!

Similar Jobs

2 Days Ago
In-Office or Remote
Senior level
Senior level
Fintech • Payments • Financial Services
Design, automate, and manage mission-critical Azure cloud infrastructure for Borrowell’s marketplace. Develop Terraform infrastructure modules, maintain Azure DevOps deployment pipelines, automate workflows with Bash, and containerize applications with Docker. Implement cloud networking and web application security, troubleshoot distributed microservices, and collaborate with development, security, and QA teams. The role is remote across eligible Canadian provinces but requires occasional in-person meetings and team events.
Top Skills: .NetAzureAzure DevopsBashCi/CdDnsDockerGithub ActionsGitlabInfrastructure As CodeKubernetesLinuxNetwork Security GroupsPrivate NetworkingRoutingSubnetsTerraformVnetsWeb Application Firewall
4 Days Ago
Remote
Canada
Senior level
Senior level
Web3
Build and operate secure, scalable, observable infrastructure for AI platforms, web applications, APIs, backend services, and shared platforms across Azure and OCI. Manage Kubernetes environments, cloud networking, CI/CD pipelines, infrastructure automation, databases, caches, queues, monitoring, incident response, reliability, and cost optimization. Translate technical designs into production-ready infrastructure and create reusable platform patterns. Collaborate with AI, application, and architecture teams while troubleshooting across application and infrastructure boundaries.
Top Skills: ArgocdAzure Kubernetes ServiceBashBitbucketC#CloudflareDockerEnvoyFastapiGoGrafanaHelmJavaJavaScriptJenkinsKubernetesLinuxLitellmAzureOpentelemetryOracle Cloud InfrastructureOracle Kubernetes EnginePostgresPrometheusPythonRabbitMQRedisSentryTerraformTypescript
10 Days Ago
Remote
Canada
Senior level
Senior level
AdTech • Marketing Tech
Own and improve AWS infrastructure, shared platform services, infrastructure-as-code workflows, Kubernetes deployments, observability, and operational reliability. Troubleshoot distributed systems, support incident response, establish infrastructure standards, and partner with engineering teams on resilient, scalable system design.
Top Skills: AnsibleArgocdAWSConsulGithub ActionsGrafanaKubernetesPackerPrometheusRedisTerraformTerrateamVault

What you need to know about the Montreal Tech Scene

With roots dating back to 1642, Montreal is often recognized for its French-inspired architecture and cobblestone streets lined with traditional shops and cafés. But what truly sets the city apart is how it blends its rich tradition with a modern edge, reflected in its evolving skyline and fast-growing tech industry. According to economic promotion agency Montréal International, the city ranks among the top in North America to invest in artificial intelligence, making it le spot idéal for job seekers who want the best of both worlds.

Key Facts About Montreal Tech

  • Number of Tech Workers: 255,000+ (2024, Tourisme Montréal)
  • Major Tech Employers: SAP, Google, Microsoft, Cisco
  • Key Industries: Artificial intelligence, machine learning, cybersecurity, cloud computing, web development
  • Funding Landscape: $1.47 billion in venture capital funding in 2024 (BetaKit)
  • Notable Investors: CIBC Innovation Banking, BDC Capital, Investissement Québec, Fonds de solidarité FTQ
  • Research Centers and Universities: McGill University, Université de Montréal, Concordia University, Mila Quebec, ÉTS Montréal

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account