Zencoder Logo

Zencoder

Senior Engineer, Infrastructure

Posted 3 Days Ago
In-Office or Remote
Hiring Remotely in Greece
Senior level
In-Office or Remote
Hiring Remotely in Greece
Senior level
Build and operate Zencoder's cloud and Kubernetes infrastructure to improve reliability, security, scalability and cost efficiency. Own GKE clusters, cloud networking/IAM, data systems (OpenSearch/Postgres), CI/CD/GitOps, observability, incident response, and automation to reduce operational toil and prepare the platform for growth.
The summary above was generated by AI
About Zencoder

At Zencoder.ai, we build and orchestrate AI agents that ship real work - code, research, operations, and more. What started as developer tooling is becoming a platform where people and agents collaborate across knowledge work tasks.

About the role

We’re looking for an Engineer to help build and operate the infrastructure behind Zencoder’s AI-powered products.

You’ll work across our production platform, improving its reliability, security, scalability and cost efficiency. This includes our Kubernetes foundations, cloud infrastructure, networking, data systems and the internal tooling that enables engineers to deploy and operate services confidently.

This is a hands-on engineering role rather than a traditional operations position. You’ll write code, automate infrastructure, investigate production issues and design systems that reduce operational complexity as the company grows.

The exact problems will evolve quickly. You should be comfortable taking ownership of unfamiliar systems, identifying the highest-leverage improvements and moving between immediate production needs and longer-term platform investments.

Example projects include
  • Owning our Kubernetes foundations: building and operating production GKE clusters with reliable networking, ingress, service-to-service communication, workload isolation, autoscaling and deployment patterns.
  • Improving cloud security and networking: evolving our GCP architecture across VPCs, IAM, workload identity, secrets, firewalls, WAF, CDN and other security controls.
  • Building dependable search infrastructure: improving the deployment, scaling, performance and operational reliability of OpenSearch and other data-intensive systems.
  • Reducing infrastructure cost: developing better cost attribution, capacity planning and optimisation across compute, storage, networking, observability and managed cloud services.
  • Making deployments safer: improving CI/CD, GitOps, progressive delivery, automated rollback and the tooling engineers use to deploy and operate their services.
  • Strengthening production reliability: improving observability, alerting, incident response, disaster recovery and the resilience of critical customer-facing systems.
  • Automating operational work: replacing manual procedures with software, infrastructure-as-code and reusable platform capabilities.
  • Preparing the platform for growth: identifying architectural bottlenecks and evolving our infrastructure to support increasing usage, larger customers and new AI workloads.
You may be a fit if
  • You have strong software-engineering skills and regularly write production code.
  • You have experience building and operating infrastructure on GCP, AWS or another major cloud platform.
  • You have hands-on experience with Kubernetes in production.
  • You understand cloud networking and security concepts such as VPCs, IAM, load balancing, firewalls, WAFs, CDNs, DNS and service identity.
  • You have experience with infrastructure-as-code and automated deployment systems.
  • You are comfortable debugging problems across application, infrastructure, networking and data-system boundaries.
  • You have operated distributed systems such as OpenSearch, Elasticsearch, PostgreSQL or similar technologies at scale.
  • Experience deploying or operating large language models with serving frameworks such as vLLM or SGLang is a plus, but not required.
  • You care about reliability, security, developer experience and cost - not just whether infrastructure is technically running.
  • You look for ways to remove operational toil rather than accepting repetitive manual work.
  • You take ownership of important problems and are comfortable working across traditional team boundaries.

Experience with every technology we use is not required. We value strong engineering fundamentals, good judgement and the ability to learn unfamiliar systems quickly.

Why Join Zencoder?
  • Shape the Future of Software Creation: We’re not just improving how developers write code — we’re redefining how ideas turn into reality. By closing the gap between concept and execution, we’re creating tools that will influence every industry that relies on software.
  • Massive Impact, Real Ownership: At Zencoder, you’ll have full visibility into how your work moves the product and the company forward. You’ll ship features that matter, see the immediate impact of your decisions, and get feedback directly from users — fast.
  • ICs Are the Core: Individual Contributors are the highest-status role at Zencoder. Our culture celebrates those who lead by doing — who create momentum, inspire others, and turn ideas into shipped products.
  • High-Caliber Team & Founder: Work alongside exceptional AI and software engineers, and learn directly from Andrew Filev, founder of a unicorn startup, who brings deep expertise in scaling world-class technology companies.
  • Global & Flexible: We hire talent, not coordinates. Work from wherever you’re happiest and most productive — as long as you bring the energy, focus, and results.
  • Aligned Incentives: Our equity plan ensures that when we succeed, you succeed. Your impact compounds as the company grows.

Zencoder is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.

Similar Jobs

Senior level
Artificial Intelligence • Cloud • Hardware • Automation
Hands-on engineer responsible for deploying, validating, and operating GPU cloud infrastructure in datacenters. Duties include rack-and-stack, cabling and optics validation, BIOS/firmware and GPU/DPU bring-up, network and storage integration, platform stack installation, acceptance testing, runbook creation, Day-2 maintenance, observability validation, and cross-team coordination with vendors and engineering for production readiness.
Top Skills: AnsibleBashBgpBluefield DpuCitrix NetscalerCloudstackCniContainerdCsiCudaDcgmDcgm ExporterDockerEcmpEvpn/VxlanGrafanaHgxIpmiKubernetesKubevirtKvmLocal NvmeLokiMellanox/Nvidia NicsMellanox/Nvidia OfedNvidia DriversNvidia Gpu OperatorNvidia NetqNvidia Spectrum/CumulusNvlinkNvmlNvswitchOob ManagementOptic TransceiversOvs/OvnPcie PassthroughPrometheusPythonQemuRdmaRedfishRoceSr-IovStorpoolTerraformUbuntu LinuxVfioVlanVrfVyosWafWekaZabbix
6 Days Ago
Remote or Hybrid
Senior level
Senior level
Artificial Intelligence • Software
As a Senior Cloud Infrastructure Engineer, you will manage Langfuse's cloud operations on AWS, ensuring high uptime and performance while optimizing costs. You'll own the observability setup with Datadog, automate processes, and make self-hosting seamless for users. You'll also focus on scaling infrastructure and maintaining security compliance.
Top Skills: Aws Ecs FargateClickhouseCloudFormationDatadogDockerExpressHelmKubernetesNext.JsPostgresPulumiRedisS3TerraformTypescript
24 Days Ago
In-Office or Remote
Senior level
Senior level
Artificial Intelligence • Software • Automation
Design, build, and operate cloud infrastructure for n8n Cloud, focusing on multi-region deployments and security, using Terraform and Kubernetes.
Top Skills: AWSAzureGCPKubernetesTerraform

What you need to know about the Montreal Tech Scene

With roots dating back to 1642, Montreal is often recognized for its French-inspired architecture and cobblestone streets lined with traditional shops and cafés. But what truly sets the city apart is how it blends its rich tradition with a modern edge, reflected in its evolving skyline and fast-growing tech industry. According to economic promotion agency Montréal International, the city ranks among the top in North America to invest in artificial intelligence, making it le spot idéal for job seekers who want the best of both worlds.

Key Facts About Montreal Tech

  • Number of Tech Workers: 255,000+ (2024, Tourisme Montréal)
  • Major Tech Employers: SAP, Google, Microsoft, Cisco
  • Key Industries: Artificial intelligence, machine learning, cybersecurity, cloud computing, web development
  • Funding Landscape: $1.47 billion in venture capital funding in 2024 (BetaKit)
  • Notable Investors: CIBC Innovation Banking, BDC Capital, Investissement Québec, Fonds de solidarité FTQ
  • Research Centers and Universities: McGill University, Université de Montréal, Concordia University, Mila Quebec, ÉTS Montréal

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account