Skip to content
Scale AIScale AI

Machine Learning Research Engineer, Agent Data Foundation - Enterprise GenAI

Develops synthetic data pipelines, production trace agents, and automated agent-building frameworks for enterprise GenAI. Requires 3+ years LLM production experience, top conference publications, and advanced CS degree.

About the job

Responsibilities

  • Build synthetic data pipelines to generate enterprise environments for RL post-training.
  • Create agents to convert production traces into actionable insights for agent improvement.
  • Contribute to agent-building product using coding agents and proprietary algorithms.
  • Train state-of-the-art models (internal and community) for enterprise deployment.

Requirements

  • 3+ years building with LLMs in production environments.
  • Experience constructing high-quality data to improve LLMs/Agents.
  • Publications in top conferences (NeurIPS, ICLR, ICML) within last two years.
  • PhD or Master's in Computer Science or related field.

Skills

LLMs, Reinforcement Learning, Synthetic Data, Agent Building, Post-Training Algorithms, Python, Machine Learning, PyTorch, Data Pipelines, Trace Analysis

Cinder

Cinder

New York, NY

AI/ML Engineer
$220k+/yrHybrid5+ YOEML Engineering

Build and operate production machine-learning systems for content safety, from messy customer data through classification, evaluation, and inference. The role requires 5+ years of ML engineering experience, strong Python and MLOps skills, and sound judgment across classical models and LLMs.

Perplexity

Perplexity

San Francisco, CA

Member of Technical Staff
$220k+/yrOn-siteML Engineering

Build AI agent harnesses, models, and product capabilities that enable agents to perform complex work across digital environments. The role combines applied AI research and software engineering, requiring Python proficiency, strong product judgment, and experience with agent tooling, reinforcement learning, or browser technologies.

Mercor

Mercor

San Francisco, CA

Member of Technical Staff, Enterprise Evals Platform
$220k+/yrOn-siteML Engineering

Builds the platform, verifiers, environments, and grading infrastructure used to evaluate enterprise AI agents at scale. The role combines strong software engineering with expertise in agent runtimes, evaluation design, benchmarks, and production failure analysis.

Applied Intuition

Applied Intuition

Sunnyvale, CA

Machine Learning Performance Engineer - Offboard Training & Inference
$215k+/yrOn-siteML Engineering

Optimizes distributed machine learning training and high-throughput offline inference across large accelerator clusters. The role focuses on profiling, scaling efficiency, cluster goodput, GPU performance, and cost-effective processing of autonomy data.

Firecrawl

Firecrawl

San Francisco, CA

Machine Learning Engineer
$210k+/yrHybrid3+ YOEML Engineering

Build and operate production machine-learning systems for search ranking, relevance, extraction quality, and LLM-driven features. The role requires production ML ownership, ranking or relevance expertise, large-scale data experience, Python, and rigorous experimentation skills.