Research Engineer
Research Engineer working on frontier audio AI models, including training, post-training, data curation, architectural improvements, and rigorous evaluation. Requires at least 3 years of AI experience and evidence of solving challenging machine learning problems through projects or research.
About the job
Responsibilities
- Train state-of-the-art models across modalities.
- Improve existing models through post-training, data curation, architectural improvements, and novel training paradigms.
- Design rigorous benchmarks and evaluations to determine whether model iterations improve and why.
Requirements
- 3+ years of experience in AI, emphasizing implementation, training, and improvement of machine learning models.
- Demonstrated ability to autonomously evaluate novel concepts or enhance existing machine learning projects, potentially contributing to published work.
- Experience conducting exploratory research to improve the quality of gathered data.
- Evidence of solving difficult problems through artifacts such as past projects, designs, or GitHub contributions.
Benefits
- Annual discretionary professional development stipend.
- Annual discretionary stipend for social travel with colleagues.
- Annual company offsite.
- Monthly co-working stipend for employees not near a main hub.
Skills
Artificial Intelligence, Machine Learning, Model Training, Post-Training, Data Curation, Model Evaluation, Benchmarking, Machine Learning Research, Data Quality, GitHub
Similar jobs
AI Research jobsConduct applied research on AI agents, designing experiments and evaluation systems to improve reliability, context retention, and multi-step task completion. The role requires strong AI/ML research, engineering, experimental design, and communication skills.
Research Engineer building large-scale AI capability evaluations, telemetry, data pipelines, and analysis tools for Anthropic’s Takeoff Intel team. The role requires hands-on large language model experimentation, rapid prototyping, data expertise, and strong research collaboration.
Conduct applied research on foundation models for fraud detection using large-scale behavioral and financial-risk data. The role spans experimentation, evaluation, production deployment, and cross-functional work on model governance, requiring 4+ years of applied ML experience and strong Python and SQL skills.
Researcher or engineer focused on designing, evaluating, and productionizing oversight systems and safety mitigations for autonomous AI agents. The role requires strong systems or security reasoning, threat-modeling ability, and experience building practical evaluations and controls.
Researcher focused on training and evaluating frontier AI agents, mining incidents, and building scalable safety measurement systems. The role requires strong research or ML engineering execution, quantitative judgment, and the ability to own ambiguous projects end to end.