Young Investigator, FlexOlmo
Postdoctoral researcher leading high-impact AI projects on Mixture-of-Experts and long-context language models, training/releasing models, building open-source tools, publishing papers, and mentoring juniors. Requires recent PhD in CS/ML with strong publication record and PyTorch expertise.
About the job
Responsibilities
- Define and lead a high-impact research project.
- Train and release leading models.
- Collaborate with and learn from team members across Ai2.
- Build open-source software for the research community.
- Author scientific papers for publication in high-profile conferences or journals.
- 50% work leading and collaborating on Ai2 project as an independent contributor.
- 50% work mentoring junior researchers (PhD students/interns, predoctoral students/interns).
Requirements
- Within one year of completing PhD or already have PhD in Computer Science or similar field.
- Research experience in machine learning, natural language processing, language and vision, or related areas.
- Outstanding individual contributor skills, especially with deep learning frameworks (e.g. PyTorch).
- Outstanding publication record at AI-related venues (NeurIPS, ICLR, ICML, COLM, ACL, EMNLP).
- Extensive research experience in large language models, training dynamics, scaling laws, and data curation.
- Experience with mixture-of-experts, long-context language models, and retrieval preferred.
- Located or willing to relocate to Berkeley, CA.
Compensation
- $159,650
Skills
PyTorch, LLMs, Mixture-Of-Experts, Long-Context Language Models, Retrieval, Natural Language Processing, Machine Learning, Scaling Laws, Data Curation, Deep Learning
Similar jobs
AI Research jobsResearch Scientist developing and evaluating health-focused AI models, large language models, and agentic systems for clinical applications. The role requires advanced research experience, strong coding skills, healthcare or clinical-data experience, and top-tier AI/ML publications.
Conduct research on long-horizon, multi-agent AI behavior by designing agent environments, analyzing large-scale data, and running experiments. The role requires strong research judgment, rapid execution, independence, and familiarity with current AI developments.
Build, optimize, and evaluate long-running and multi-agent AI systems, along with tools for monitoring and analyzing their real-world behavior. The role requires software engineering experience with coding agents, strong independence, and familiarity with current AI developments.
Develops experimental AI techniques and prototypes for agentic marketing applications, with emphasis on image and video generation. The role requires strong backend or probabilistic systems expertise, quantitative thinking, creativity with LLM applications, and product intuition.
The Research Engineer will apply advances in agents and language models to build and evaluate multi-agent systems for automated code validation and review. The role requires a computer science or equivalent background, research experience, strong programming skills, and product intuition.