Research Staff, LLMs
Conduct cutting-edge research on Large Language Models, focusing on transformer optimization, distributed training, data curation, and RL. Collaborate on experiments, deploy models to production, and drive voice AI innovations.
About the job
What You'll Do
- Brainstorming and collaborating with other members of the Research Staff to define new LLM research initiatives
- Broad surveying of literature, evaluating, classifying, and distilling current methods
- Designing and carrying out experimental programs for LLMs
- Driving transformer (LLM) training jobs successfully on distributed compute infrastructure and deploying new models into production
- Documenting and presenting results and complex technical concepts clearly for a target audience
- Staying up to date with the latest advances in deep learning and LLMs, with a particular eye towards their implications and applications within our products
It's Important to Us That You Have
- 3+ years of experience in applied deep learning research, with a solid understanding toward the applications and implications of different neural network types, architectures, and loss mechanism
- Proven experience working with large language models (LLMs) - including experience with data curation, distributed large-scale training, optimization of transformer architecture, and RL Learning
- Strong experience coding in Python and working with Pytorch
- Experience with various transformer architectures (auto-regressive, sequence-to-sequence.etc)
- Experience with distributed computing and large-scale data processing
- Prior experience in conducting experimental programs and using results to optimize models
It Would Be Great if You Had
- Deep understanding of transformers, causal LMs, and their underlying architecture
- Understanding of distributed training and distributed inference schemes for LLMs
- Familiarity with RLHF labeling and training pipelines
- Up-to-date knowledge of recent LLM techniques and developments
- Published papers in Deep Learning Research, particularly related to LLMs and deep neural networks
Benefits & Perks
- Medical, dental, vision benefits
- Annual wellness stipend
- Mental health support
- Unlimited PTO
- Generous paid parental leave
- Flexible schedule
- 401(k) plan with company match
- Learning / Education stipend
Skills
LLMs, Transformers, PyTorch, Python, Distributed Training, Reinforcement Learning, Data Curation, RLHF, Causal Language Models, Deep Learning
Similar jobs
AI Research jobsEvaluates model and Generative AI risks across Upstart Bank’s model inventory, conducting risk assessments, monitoring reviews, quantitative analyses, and governance activities. Requires a quantitative master’s degree, 4+ years of relevant experience, and coding skills in Python, R, or similar languages.
Leads technical direction and develops maritime autonomy capabilities for unmanned surface and underwater vehicles, including motion planning, localization, safe behaviors, and heterogeneous multi-agent collaboration. Requires deep robotics and unmanned-systems experience, strong C++/Python skills, and senior technical leadership.
Leads design, implementation, integration, and field validation of tactical autonomy software for unmanned systems and multi-agent missions. Requires extensive autonomy or robotics experience, strong C++ and Python skills, and the ability to obtain a SECRET clearance.
Research and evaluate frontier AI capabilities for cybersecurity, rapidly prototyping tools, designing rigorous benchmarks, and helping operationalize reliable capabilities into products. Requires deep security expertise, strong technical communication, and at least seven years of relevant experience.
Own the architecture, delivery, evaluation, and production operations of AI capabilities embedded in procurement and finance workflows. The role requires 10+ years in applied AI or machine learning, deep LLM and agent expertise, and experience delivering measurable production outcomes.