Skip to content
LatentLatent

Machine Learning Engineer

Owns end-to-end production ML systems for clinical workflows, including training/fine-tuning LLMs for medical reasoning and question answering. Requires strong ML/software engineering, PyTorch experience, and ability to handle high-stakes ambiguity with real patient impact.

About the job

What You’ll Do

  • Own end-to-end ML systems, including architecture, data, modeling, evaluation, and production infrastructure
  • Train and fine-tune large language models (LLMs) for:
    • Clinical reasoning
    • Medical question answering
    • Evidence-grounded generation
  • Make and own tradeoffs across accuracy, latency, cost, and safety in high-stakes production environments
  • Develop evaluation frameworks to ensure model safety and clinical validity
  • Integrate ML systems into product workflows and patient-facing applications
  • Monitor system performance in production and iterate based on real-world usage and feedback
  • Define what “correct” means in ambiguous clinical workflows in collaboration with engineers and clinicians

What We’re Looking For

  • Strong foundation in machine learning and software engineering
  • Track record of building and owning ML systems in production where performance, reliability, or correctness materially mattered
  • Experience driving ambiguous ML problems from 0→1, including problem formulation, model design, and productionization
  • Hands-on experience with PyTorch or similar frameworks
  • Ability to operate independently in high-ambiguity environments with minimal guidance
  • Strong product and engineering judgment — you know when to use ML, when not to, and how to scope problems accordingly
  • Comfort working in a fast-moving, early-stage environment
  • Experience working on systems where decisions have real-world consequences (e.g., healthcare, finance, infrastructure)

Nice to Have

  • Experience deploying LLMs in production environments
  • Experience building distributed systems or large-scale data pipelines
  • Experience working with clinical, biomedical, or other regulated datasets

Compensation

Base salary: $225,000 – $300,000+ Meaningful equity in an early-stage, Series A company

Skills

PyTorch, LLMs, Reinforcement Learning, Machine Learning, Llm Fine-Tuning, Distributed Systems, Data Pipelines, Clinical Data, Production Ml, Evaluation Frameworks

Cinder

Cinder

New York, NY

AI/ML Engineer
$220k+/yrHybrid5+ YOEML Engineering

Build and operate production machine-learning systems for content safety, from messy customer data through classification, evaluation, and inference. The role requires 5+ years of ML engineering experience, strong Python and MLOps skills, and sound judgment across classical models and LLMs.

Perplexity

Perplexity

San Francisco, CA

Member of Technical Staff
$220k+/yrOn-siteML Engineering

Build AI agent harnesses, models, and product capabilities that enable agents to perform complex work across digital environments. The role combines applied AI research and software engineering, requiring Python proficiency, strong product judgment, and experience with agent tooling, reinforcement learning, or browser technologies.

Mercor

Mercor

San Francisco, CA

Member of Technical Staff, Enterprise Evals Platform
$220k+/yrOn-siteML Engineering

Builds the platform, verifiers, environments, and grading infrastructure used to evaluate enterprise AI agents at scale. The role combines strong software engineering with expertise in agent runtimes, evaluation design, benchmarks, and production failure analysis.

Applied Intuition

Applied Intuition

Sunnyvale, CA

Machine Learning Performance Engineer - Offboard Training & Inference
$215k+/yrOn-siteML Engineering

Optimizes distributed machine learning training and high-throughput offline inference across large accelerator clusters. The role focuses on profiling, scaling efficiency, cluster goodput, GPU performance, and cost-effective processing of autonomy data.

Garner Health

Garner Health

New York, NY

Applied Scientist III
$236k+/yrOn-site4+ YOEML Engineering

Build and deploy algorithmic systems for high-impact healthcare problems, choosing among machine learning, optimization, heuristics, and hybrid approaches. The role requires 4+ years of relevant industry experience, strong applied problem-solving and evaluation skills, and fluency in modern ML tooling.