Norbert Health Logo

Norbert Health

Applied AI Engineer

Reposted One Month Ago
Be an Early Applicant
In-Office
Montréal, QC, CAN
Mid level
In-Office
Montréal, QC, CAN
Mid level
Develop automated, production-grade MLOps pipelines for AI in healthcare, focusing on integrating foundation models and real-time streaming systems under regulatory constraints.
The summary above was generated by AI
The company

Norbert is building autonomous robots that deliver healthcare.

Our AI sensing platform enables existing robotic platforms to become care team members: rounding on patients, capturing vitals without contact (FDA-cleared for pulse and respiratory rate, more in the pipeline), running assessments, documenting to the EMR, and escalating when something’s wrong. Autonomously.

We’re not building demos. We’re deployed in real facilities today, monitoring hundreds of patients daily. We’re solving one of healthcare’s hardest problems: a global nursing shortage that will hit 40% by 2030.

We’re a small, international team backed by top-tier VCs, with offices in Brooklyn, Paris, and Montreal. We ship things that matter.

The position

We're looking for an Applied AI Engineer to take our growing collection of foundation models and ML components from manually run, sometimes locally trained workflows to fully automated, production-grade MLOps pipelines: deployed reliably on robots in nursing facilities.  We need someone who knows the model landscape cold, treats evaluation as a first-class engineering problem, and has strong opinions about when to prompt, RAG, fine-tune, swap, or buy.

You’ll work across cloud and edge deployments, and some of the systems you’ll touch are on a SaMD pathway, so you’ll need to be comfortable shipping under regulatory constraints.

What you’ll do
  • Integrate foundation models and ML components (VLMs, LLMs, ASR/TTS, detection/segmentation, embeddings) into our production pipelines, using both open-weight models and third-party APIs
  • Build RAG and agent-style orchestration for clinical reporting and conversational interfaces
  • Ship real-time streaming pipelines (voice agents) alongside batch and request-response workloads
  • Build evaluation harnesses that catch regressions across model swaps and measure performance against clinical-grade accuracy targets
  • Fine-tune and retrain models (LoRA, PEFT, supervised fine-tuning) using data collected from our deployed fleet
  • Deploy across our inference surfaces: third-party APIs, self-hosted, and on-robot edge
  • Build the data flywheel: pipelines that collect, label, version, and feed production data back into model improvement
  • Partner with the algorithms team (signal processing, computer vision) on integration with their lower-level pipelines
What we’re looking for
  • BS in Computer Science, Engineering, or a related field, or equivalent hands-on experience
  • 4+ years shipping ML/AI systems in production outside of academic settings
  • Strong working knowledge of the modern foundation model landscape (open-weight LLMs and VLMs, common detection/segmentation backbones, embedding models)
  • Hands-on experience with PEFT/LoRA and supervised fine-tuning
  • Strong Python; comfortable with the deployment toolchain (ONNX, quantization, at least one inference runtime—TensorRT, vLLM, llama.cpp, etc.)
  • Experience with a cloud ML training/MLOps platform (GCP Vertex AI, AWS SageMaker, Azure ML, or equivalent)
  • Ability to work independently, solve complex problems, and drive projects to completion
Bonus points
  • Edge ML deployment (Jetson, ARM, mobile NPUs)
  • Real-time voice AI pipelines (STT, TTS, streaming LLM)
  • Production RAG systems beyond toy implementations
  • Medical devices, SaMD, or other regulated ML environments
  • MLOps tooling (Weights & Biases, MLflow, DVC, etc.)
  • Active learning or human-in-the-loop labeling workflows
  • C++ for integrating with our computer vision pipeline
What we offer
  • Real impact: your code provides care for patients today
  • High autonomy and technical ownership—you’ll define how we operate AI in production
  • Work at the intersection of cutting-edge AI, edge computing, and healthcare
  • A talented, excellent, diverse and international team
  • Equity participation in the company’s future
  • Cutting-edge stack: embedded AI, robotics, LLMs, multimodal sensing
  • Transparent, mission-driven culture focused on continuous learning
  • Competitive salary and equity

Similar Jobs

8 Days Ago
In-Office or Remote
CA
Senior level
Senior level
Logistics • Transportation
Architect, build, deploy, and operate production-grade AI agentic systems and workflows. Responsibilities include backend and full-stack development, agent orchestration, workflow state, retrieval, memory, approvals, failure recovery, data architecture, enterprise integrations, cloud infrastructure, observability, security, and scaling. The role establishes reusable engineering patterns and partners with users to deliver reliable AI products, while addressing production issues involving model behavior, permissions, latency, cost, and consistency.
Top Skills: Ci/CdContainersEvent-Driven ArchitecturesGoJavaScriptLangchainLanggraphLlamaindexMcpPythonRagRelational DatabasesTemporalTypescriptVector Search
16 Days Ago
Remote or Hybrid
CA
Mid level
Mid level
Artificial Intelligence • Information Technology • Software
Partners with AI labs to integrate models, run evaluations, deliver results, and manage customer relationships. Builds pragmatic full-stack solutions, automations, onboarding workflows, and delivery tooling for cutting-edge AI teams. Translates customer and research feedback into product and engineering initiatives, supports strategic account growth, and serves as a technical point of contact across onboarding, releases, and issue resolution.
Top Skills: Ci/CdCloud InfrastructureLlm Provider ApisTypescript
29 Days Ago
Remote or Hybrid
Canada
Senior level
Senior level
Big Data • Information Technology • Software • Database • Analytics
Develop experimental AI techniques for agentic marketing applications, especially image and video generation. Build proofs of concept, autonomous evaluation and improvement systems, realistic AI-generated video, and brand-aligned creative generation workflows. The role requires strong quantitative and probabilistic thinking, backend architecture skills, creativity with LLM applications, and product intuition.
Top Skills: Agentic AiBackend ArchitectureGenerative AiImage GenerationLlmsMachine LearningProbabilistic SystemsVideo Generation

What you need to know about the Montreal Tech Scene

With roots dating back to 1642, Montreal is often recognized for its French-inspired architecture and cobblestone streets lined with traditional shops and cafés. But what truly sets the city apart is how it blends its rich tradition with a modern edge, reflected in its evolving skyline and fast-growing tech industry. According to economic promotion agency Montréal International, the city ranks among the top in North America to invest in artificial intelligence, making it le spot idéal for job seekers who want the best of both worlds.

Key Facts About Montreal Tech

  • Number of Tech Workers: 255,000+ (2024, Tourisme Montréal)
  • Major Tech Employers: SAP, Google, Microsoft, Cisco
  • Key Industries: Artificial intelligence, machine learning, cybersecurity, cloud computing, web development
  • Funding Landscape: $1.47 billion in venture capital funding in 2024 (BetaKit)
  • Notable Investors: CIBC Innovation Banking, BDC Capital, Investissement Québec, Fonds de solidarité FTQ
  • Research Centers and Universities: McGill University, Université de Montréal, Concordia University, Mila Quebec, ÉTS Montréal

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account