Research Engineer - Mid-Training

Trains frontier LLMs on semiconductor design/verification data (RTL, netlists, PDKs) for automated chip development. Develops synthetic data generation, model distillation, evals, and scales training across thousands of GPUs.

Palo Alto, CAML EngineeringOnsite

Apply

About the role

Responsibilities

Train frontier models to become highly knowledgeable semiconductor design and verification experts for reinforcement learning and automated chip development.
Develop methods for generating and curating synthetic design data, performing model distillation, and enabling continual learning at scale.
Work with hardware engineers, RL researchers, and verification specialists to create evals that guide design data quality and model improvement.
Collaborate with compute engineers to scale efficient training across thousands of GPUs and RL environments.
Build high-performance tools to investigate how data and simulation shape model-driven design intelligence.

Requirements

Experience training LLMs or foundation models on semiconductor design and verification corpora (e.g., RTL, netlists, PDKs, simulation logs).
Modeling design scaling laws and optimizing compute budgets for chip-design-specific workloads.
Generating large-scale synthetic design data (e.g., RTL variants, testbenches, verification traces).
Building evals that correlate with downstream design metrics (e.g., timing closure, power, area, verification coverage).

Skills

LLMsFoundation ModelsRtlNetlistsPdksSimulation LogsScaling LawsSynthetic DataTestbenchesVerification TracesEvalsTiming ClosurePyTorchGpu TrainingReinforcement Learning

Similar roles

ML Engineering jobs

Mirage

ML Engineer, Agentic Systems

ML Engineer building and improving agentic systems powered by LLMs for multimodal video understanding, reasoning, and creative editing tasks at an AI-native video platform. Requires strong production ML experience with transformers, fine-tuning, and experimental rigor.

175k – 275kNew York, NYML EngineeringOn-siteLLMsPython

Mirage

Software Engineer, Agents

Design and build agentic systems for AI-native video creation, integrating LLMs and evaluation frameworks to power creative workflows. Requires 5+ years building ML/agentic systems in production.

175k – 275kNew York, NYML EngineeringOn-site5+ YOERAGLLMs

Pindrop

Research Scientist II

Research Scientist II building and improving fraud risk models and scam detection systems using audio, behavioral, and metadata signals. Requires an advanced degree and 3+ years of applied ML experience with Python and modern ML frameworks.

160k – 185kUnited StatesML EngineeringRemote3+ YOELLMsKeras

Harvey

Research Engineer, Post-Training

Research engineer focused on post-training LLMs and agents for legal work. Requires hands-on experience training open-weight models and strong Python/research engineering skills.

231k – 340kSan Francisco, CAML EngineeringHybridSftRLHF

AI Fund

AI Engineer

Build full-stack AI prototypes and agentic systems to pressure-test venture ideas. Requires 3+ years building production AI applications with strong frontend/backend fluency and frontier coding agent expertise.

150k – 190kMountain View, CAML EngineeringOn-site3+ YOESQLAPIs