Research Engineer - Post-Training

Post-trains frontier AI models using reinforcement learning to autonomously handle semiconductor design tasks like chip architecture optimization, RTL code generation, simulations, and verification. Collaborates with hardware experts to build RL environments, reward functions, and evaluation frameworks.

Palo Alto, CAML EngineeringOnsite

Apply

About the role

Responsibilities

Post-train frontier models to autonomously perform complex tasks across the semiconductor design and verification pipeline.
Propose and optimize chip architectures.
Generate and refine RTL code.
Run simulations.
Identify verification gaps and iteratively improve designs.
Collaborate with hardware design, verification, and computer architecture experts to design reinforcement learning environments.
Develop structured reward functions, scaling strategies, and evaluation frameworks.

Requirements

Experience creating and scaling RL environments for LLMs or multimodal agents.
Building high-quality evaluation datasets and benchmarks for complex reasoning or design tasks.
Working closely with domain experts in hardware and verification to define evaluation metrics, constraints, and simulation conditions.
Designing reward functions and feedback pipelines that balance correctness, performance, and design efficiency.
Running large-scale RL fine-tuning or post-training experiments for frontier models.
Applying reinforcement learning or curriculum learning to structured reasoning or symbolic domains.

Skills

Reinforcement LearningLLMsMultimodal AgentsRl EnvironmentsEvaluation DatasetsBenchmarksRtlHardware DesignVerificationChip ArchitectureReward FunctionsRl Fine-TuningCurriculum LearningSemiconductor Design

Similar roles

ML Engineering jobs

Mirage

ML Engineer, Agentic Systems

ML Engineer building and improving agentic systems powered by LLMs for multimodal video understanding, reasoning, and creative editing tasks at an AI-native video platform. Requires strong production ML experience with transformers, fine-tuning, and experimental rigor.

175k – 275kNew York, NYML EngineeringOn-siteLLMsPython

Mirage

Software Engineer, Agents

Design and build agentic systems for AI-native video creation, integrating LLMs and evaluation frameworks to power creative workflows. Requires 5+ years building ML/agentic systems in production.

175k – 275kNew York, NYML EngineeringOn-site5+ YOERAGLLMs

Pindrop

Research Scientist II

Research Scientist II building and improving fraud risk models and scam detection systems using audio, behavioral, and metadata signals. Requires an advanced degree and 3+ years of applied ML experience with Python and modern ML frameworks.

160k – 185kUnited StatesML EngineeringRemote3+ YOELLMsKeras

Harvey

Research Engineer, Post-Training

Research engineer focused on post-training LLMs and agents for legal work. Requires hands-on experience training open-weight models and strong Python/research engineering skills.

231k – 340kSan Francisco, CAML EngineeringHybridSftRLHF

AI Fund

AI Engineer

Build full-stack AI prototypes and agentic systems to pressure-test venture ideas. Requires 3+ years building production AI applications with strong frontend/backend fluency and frontier coding agent expertise.

150k – 190kMountain View, CAML EngineeringOn-site3+ YOESQLAPIs