Research Engineer/Research Scientist, Personal AGI-Model Experience
Research Engineer/Scientist on OpenAI's Personal AGI Model Experience team, shaping ChatGPT's character, behavior, and human-AI interactions through research, human data, evaluations, reward models, and post-training. Requires strong ML engineering and research experience with large models.
About the job
Responsibilities
- Own and pursue a research agenda to improve model capability and performance.
- Collaborate closely with other research and product teams.
- Build robust evaluations for tracking modeling improvements.
- Design, implement, test, and debug code across the research stack.
Requirements
- Deep understanding of machine learning and machine learning applications.
- Good judgment about model behavior and ability to communicate this judgment effectively.
- Enjoy taking ambitious, qualitative problems and turning them into concrete training interventions.
- Working knowledge of relevant models and building evaluations for model capability improvement.
- Comfortable diving into a large ML codebase to debug.
- Thrive in a dynamic, technically complex, and collaborative environment.
- Strong ML engineering skills and research experience, especially with novel and highly capable models.
- Passionate about product-driven research and the quality of human-AI interaction.
Nice-to-Haves
- Experience in research areas combining reinforcement learning and products.
Skills
Machine Learning, Reinforcement Learning, Python, Evaluations, Reward Models, Post-Training, Ml Engineering, Debugging, Research
Similar jobs
ML Engineering jobsBuild and optimize OpenAI’s inference stack for AWS Trainium across high-performance kernels, compilers, runtimes, and model execution. The role requires systems programming and accelerator experience, with opportunities to solve end-to-end performance problems for frontier-scale AI models.
Build and deploy LLM-powered tools, agents, and ecosystem infrastructure with life sciences research institutions. The role requires deep scientific or biomedical research experience, production software development expertise, and the ability to translate partner workflows into scalable AI systems.
Build research infrastructure and tooling that enables AI models to design silicon, including reinforcement learning environments, EDA integrations, evaluations, and experiment workflows. The role requires strong software engineering fundamentals and comfort working across research, tooling, and chip-design systems.
Build and optimize the production LLM inference runtime for frontier models on OpenAI’s custom silicon. The role spans scheduling, distributed execution, memory and KV-cache management, performance tooling, and hardware-software co-design.
Build and operate machine learning models for sales roleplay, scoring, and coaching products, owning the lifecycle from fine-tuning and evaluation through production and on-device deployment. The role emphasizes open-source models, latency and privacy optimization, and rigorous model testing.