Skip to content
AfterQueryAfterQuery

Environment Engineers

Designs datasets, evaluation rubrics, and reward signals for RLHF/RLVR pipelines to expose model failure modes and improve frontier AI capabilities. Partners with AI labs; requires 1-4 YOE and passion for data-driven model behavior.

About the job

What You'll Do

  • Design data slices and explore data shapes that expose meaningful model failure modes across domains like finance, code, and enterprise workflows
  • Build and refine evaluation rubrics and reward signals for RLHF and RLVR training pipelines
  • Model annotator behavior and run experiments to improve different model capabilities
  • Develop quantitative frameworks for measuring dataset quality, diversity, and downstream impact on model alignment and capability
  • Create and manage both real world & synthetic data pipelines
  • Partner with lab research teams to translate their training objectives into concrete data and evaluation specifications

What We're Looking For

  • 1-4 YOE
  • Major plus if they've worked for/interned for any RL environment companies in the past or any AI safety or benchmarking orgs like METR, Artificial Analysis, etc.
  • Genuine obsession with how data structure, selection, and quality drive model behavior
  • Ability to design lightweight experiments, move fast, and extract actionable insights from messy results
  • Former founders and early engineers at early stage startups are a plus. We don't filter on pedigree. We want people who can demonstrate they work hard, learn fast, and care deeply about getting the details right.

Compensation

$200k base + profit share (around 150% of base) + competitive equity

Skills

RLHF, Rlvr, Data Pipelines, Synthetic Data, Evaluation Frameworks, Dataset Quality Metrics, Ai Benchmarking, Model Alignment, Data Slices, Quantitative Analysis

SentiLink

SentiLink

United States

Applied ML Scientist, New Grad
$180k+/yrRemoteML Engineering

Build and deploy production machine-learning models for fraud detection, identity verification, and financial risk products. The role suits new PhD graduates or early-career researchers with strong quantitative foundations, Python experience, and interest in owning the full ML lifecycle.

Earnin

Earnin

Mountain View, CA

Machine Learning Engineer
$187k+/yrHybrid2+ YOEML Engineering

Machine learning engineer who trains, evaluates, and productionizes models and LLM-powered applications for financial products. Requires 2+ years of ML systems experience, strong Python and PyTorch skills, production data pipelines, model evaluation, and API development.

Ambral

Ambral

New York, NY
Member of Technical Staff
$165k+/yrOn-site1+ YOEML Engineering

Build replayable enterprise environments, evaluation systems, graders, and post-training workflows for AI agents. The role spans machine-learning research and production engineering and requires 1–7 years of software or ML systems experience.

Nuro

Nuro

Mountain View, CA

Software Engineer, ML Inference Platform
$160k+/yrOn-site1+ YOEML Engineering

Build and maintain machine learning infrastructure for autonomy teams, including model pipelines, observability, inference serving, and compiler platforms. The role requires a relevant degree, at least one year of experience, strong Python skills, and familiarity with C++.

Nuro

Nuro

Mountain View, CA

Software Engineer, ML Infrastructure Platform
$160k+/yrOn-site1+ YOEML Engineering

Build and operate the infrastructure powering large-scale machine-learning training for autonomous-driving systems. The role requires Python proficiency, Kubernetes production experience, distributed-systems expertise, and ownership of reliability, observability, and operational maturity.