LawZero Logo

LawZero

Platform Engineer

Posted 10 Days Ago
Be an Early Applicant
In-Office
Montréal, QC, CAN
Mid level
In-Office
Montréal, QC, CAN
Mid level
Own and operate the platform supporting AI research workloads, including Kubernetes clusters, CI/CD pipelines, cloud infrastructure, internal services, observability, and security controls. Manage infrastructure as code, connect Kubernetes with HPC and Slurm environments, establish platform standards, document systems, automate recurring work, and participate in incident response. The role requires close collaboration with researchers while maintaining secure, reliable, and usable infrastructure.
The summary above was generated by AI

Founded by Yoshua Bengio, LawZero is a nonprofit organization focused on AI safety. In charge of technology, the IT department oversees Cybersecurity, End User support, and the management of the compute environment used to achieve our mission.

You will own the platform layer that sits between our people and the GPU infrastructure: CI/CD pipelines, Kubernetes clusters, cloud environments, and the standards that keep them consistent and secure. This is a hands-on role on a small IT team, with a wide surface area and a lot of room to shape how things are built.

Key responsibilities

  • Design, deploy, and run Kubernetes clusters for research workloads, including autoscaling, network policies, and workload isolation, both on-premise and in the cloud.
  • Define, communicate, and enforce best practices for CI/CD pipelines: automated builds, tests, container image creation, and deployments, with security scanning and provenance built into the pipeline rather than bolted on.
  • Implement and manage internal services supporting our research and data teams, providing them with a reliable, secure platform. Examples include artifacts registry, Container registries, etc.
  • Manage our cloud environments as code (Terraform or equivalent): accounts, networking, identity, secrets, and cost visibility.
  • Define and champion platform standards. Base images, deployment patterns, environment promotion, so researchers and engineers ship without reinventing the plumbing each time.
  • Work at the boundary between Kubernetes and our HPC cluster: containerized workflows that need to interface with the ones running on Slurm, shared storage access, and tooling that makes both environments feel coherent to a researcher. 
  • Build observability into the platform: metrics, logs, traces, and alerts that make failures obvious and debugging quick. We currently use Prometheus and Grafana.
  • Embed security controls into everything above: least-privilege IAM, secrets management, supply-chain integrity for dependencies and images, network segmentation, and audit trails.
  • Document what you build and automate what you repeat. Reduce the number of things that only work because one person remembers how.
  • Participate in incident response for platform services, and in the post-incident work that keeps the same problem from recurring.

Skills and qualifications

  • Experience

    • 3–5 years in platform engineering, DevOps, SRE, or a closely related infrastructure role.
    • Production experience with Kubernetes. Not just deploying to it, but operating it: upgrades, RBAC, networking, storage, troubleshooting a cluster that is misbehaving.
    • Solid CI/CD experience with a modern toolchain (GitHub Actions, GitLab CI, Jenkins, or similar), including building pipelines from scratch. 
    • Experience architecting, deploying, and maintaining a GitOps workflow is an asset.
    • Hands-on experience with at least one major cloud provider (AWS, GCP, or Azure) and infrastructure as code. Familiarity with on-premise infrastructure is a plus.
    • Strong Linux fundamentals and comfort with a scripting or programming language such as Python, Go, or Bash.
    • Working knowledge of containers beyond the basics: image layering, registries, runtime security, minimal base images.
    • Fluency in written and spoken English, French is a strong asset.
    • Experience supporting ML or research workloads: GPU scheduling, distributed training, large datasets, high-throughput storage is an asset.

    Security mindset

    This matters as much as the technical checklist. We are looking for someone who:

    • Thinks about the blast radius of a change before making it, and about who could abuse an access path that was opened for convenience.
    • Treats secrets, credentials, and access as first-class design concerns rather than afterthoughts.
    • Understands supply-chain risk in a build pipeline: dependency provenance, image signing, artifact integrity, what a compromised runner could reach.
    • Applies least privilege by default and can explain to a researcher why a control exists, without being obstructive about it.
    • Has practical familiarity with identity and access management, network segmentation, and secure defaults in cloud environments.

    Ways of working

    • Comfortable being the person who owns a domain end to end in a small team, without a large org to hand things off to.
    • Able to work with researchers whose priorities are speed and flexibility, and find solutions that are both secure and genuinely usable.
    • Clear written communication. You will write documentation and design notes that others depend on.
    • Capable of managing multiple priorities and adjusting to a frequently changing environment.

What we offer

  • The opportunity to contribute to a unique mission with a major impact
  • Comprehensive health benefits
  • A minimum of 20 days vacation per year upon start
  • A minimum retirement savings employer contribution of 4%
  • Generous flexible benefits designed to contribute to your well-being
  • A team of passionate experts in their field
  • A collaborative and inclusive work environment with offices in the heart of Little Italy, in the trendy Mile-Ex district, close to public transportation

About LawZero

LawZero is a non-profit organization committed to advancing research and creating technical solutions that enable safe-by-design AI systems. Its scientific direction is based on new research and methods proposed by Professor Yoshua Bengio, the most cited AI researcher in the world. Based in Montreal, LawZero’s research aims to build non-agentic AI that learns primarily to understand the world rather than to act in it, giving truthful answers to questions based on transparent and externalized probabilistic reasoning. Such AI systems could be used to accelerate scientific discovery, to provide oversight for agentic AI systems, and to advance the understanding of AI risks and how to avoid them. LawZero believes that AI should be cultivated as a global public good—developed and used safely towards human flourishing. For more information, visit www.lawzero.org

You belong here

At LawZero, diversity is important to us. We value a work environment that is fair, open and respectful of differences. We welcome applications from highly qualified individuals interested in working towards our mission in a respectful, inclusive and collaborative setting.

Your personal information will be collected and processed by LawZero to evaluate your application for employment in compliance with our Privacy Policy. Under privacy laws in force in your country of residence, you may have several privacy rights, such as to request access to your personal information or to request that your personal information be rectified or erased. Details on how you can exercise your rights can be found in our Privacy Policy.

Similar Jobs

8 Days Ago
In-Office or Remote
CA
Entry level
Entry level
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Build and operate highly reliable card-issuing software and tooling across authorization, clearing, settlement, certification, and network integrations. Serve as a technical expert on card-network specifications, protocols, mandates, and production behavior. Lead complex cross-team engineering work, debug protocol-level issues, influence network requirements, mentor engineers, improve platform reliability, and participate in on-call support for critical financial systems.
Top Skills: Ai-Assisted Engineering ToolsAmerican ExpressCard NetworksIso 8583MastercardVisa
9 Days Ago
Remote or Hybrid
CA
Expert/Leader
Expert/Leader
Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Staff Software Engineer responsible for building and operating Block’s card issuing platform at the network boundary. The role requires deep expertise in card networks, ISO 8583, authorization and clearing flows, certifications, mandates, and direct network connectivity. Responsibilities include writing production code, developing conformance tooling, debugging protocols, leading cross-team technical initiatives, influencing network requirements, mentoring engineers, supporting AI-assisted development, and participating in on-call operations for critical financial systems.
Top Skills: 3-D SecureAi-Assisted Engineering ToolsAmerican ExpressCard Network ConnectivityClearing And SettlementDigital Wallet ProvisioningIso 8583MastercardPayment AuthorizationVisa
5 Days Ago
In-Office or Remote
Canada
Senior level
Senior level
Database
Design, build, and operate Supabase’s global edge and networking infrastructure, including routing, load balancing, DNS, WAF, TLS termination, CDN integrations, and observability. Improve latency, reliability, security, and cost efficiency across cloud regions. The role includes developing automated routing systems and CI/CD, troubleshooting distributed systems, participating in on-call and incident response, and collaborating with infrastructure and product teams.
Top Skills: AnycastAWSAzureBgpCdkCdnCi/CdDnsEnvoyGCPGlobal RoutingGoHaproxyHTTPIpv6Load BalancingNginxPulumiTcp/IpTerraformTlsTraefikWaf

What you need to know about the Montreal Tech Scene

With roots dating back to 1642, Montreal is often recognized for its French-inspired architecture and cobblestone streets lined with traditional shops and cafés. But what truly sets the city apart is how it blends its rich tradition with a modern edge, reflected in its evolving skyline and fast-growing tech industry. According to economic promotion agency Montréal International, the city ranks among the top in North America to invest in artificial intelligence, making it le spot idéal for job seekers who want the best of both worlds.

Key Facts About Montreal Tech

  • Number of Tech Workers: 255,000+ (2024, Tourisme Montréal)
  • Major Tech Employers: SAP, Google, Microsoft, Cisco
  • Key Industries: Artificial intelligence, machine learning, cybersecurity, cloud computing, web development
  • Funding Landscape: $1.47 billion in venture capital funding in 2024 (BetaKit)
  • Notable Investors: CIBC Innovation Banking, BDC Capital, Investissement Québec, Fonds de solidarité FTQ
  • Research Centers and Universities: McGill University, Université de Montréal, Concordia University, Mila Quebec, ÉTS Montréal

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account