Research Engineer / Research Scientist, Post-Training
Research and develop improvements to pre-trained models for deployment in ChatGPT and API using reinforcement learning and product-driven approaches. Requires strong ML engineering, research experience with novel models, and ability to debug large codebases.
About the job
In this role, you will:
- Own and pursue a research agenda to improve model capability and performance.
- Collaborate closely with the other research and product teams, allowing customers to optimize their own models.
- Build robust evaluations for tracking modeling improvements.
- Design, implement, test, and debug code across our research stack.
You might thrive in this role if you:
- Have a deep understanding of machine learning and machine learning applications.
- Have a working knowledge of relevant models, and building evaluations for model capability improvement.
- Are comfortable diving into a large ML codebase to debug.
- Thrive in a dynamic and technically complex environment.
Skills
Machine Learning, Reinforcement Learning, PyTorch, TensorFlow, Model Evaluation, Ml Engineering, Research, Python, Deep Learning, LLMs
Similar jobs
AI Research jobsConducts frontier AI research for health, developing and evaluating scalable training methods, models, and agents that improve medical reasoning, reliability, and real-world outcomes. Requires exceptional machine learning or biomedical AI research depth, hands-on coding and experimentation, and end-to-end ownership of ambiguous problems.
Conducts hands-on medicinal chemistry research to evaluate AI-generated molecules and synthetic routes, advancing small-molecule programs from design through experimental validation. Requires a chemistry PhD, sustained synthetic experience, and cross-functional collaboration skills.
Applied research scientists develop deep-learning and generative media systems for video, audio, and multimodal editing features that ship to millions of users. The role requires strong PyTorch or TensorFlow skills, rapid experimentation, and evidence of impactful research or production machine-learning work.
Conducts causal inference research for financial market prediction and portfolio optimization, developing and validating models from research through live trading. Requires Ph.D.-level coursework, strong causal inference and statistics expertise, mathematical ability, and production Python skills.
Research Engineer building large-scale AI capability evaluations, telemetry, data pipelines, and analysis tools for Anthropic’s Takeoff Intel team. The role requires hands-on large language model experimentation, rapid prototyping, data expertise, and strong research collaboration.