Staff Research Scientist, AI Agents & LLMs
Leads research in agentic AI and LLMs, developing models for enterprise reasoning, autonomous agents with tool use, and production systems. Requires PhD, expertise in LLM training/fine-tuning, agent systems, and technical leadership.
About the job
Responsibilities
- Define research direction in Agentic AI and LLMs
- Develop models: train, fine-tune, and align models for enterprise reasoning and tool use
- Advance autonomous agents: drive state-of-the-art capabilities in multi-step reasoning and tool use
- Advance the stack: retrieval, grounding, memory, and multi-agent coordination
- Establish evaluation standards for reliability, safety, and efficiency
- Drive research breakthroughs into production within Snowflake’s platform
- Mentor and lead to amplify the team’s technical impact
- Contribute to the field through publications and open source
Requirements
- Ph.D. in Computer Science or a related field with a strong publication record
- Expertise in LLM development (training, fine-tuning, alignment, or post-training)
- Experience with agentic systems (multi-agent systems, tool use, or agent optimization)
- Strong systems thinking under real-world constraints (latency, cost, reliability)
- Proven technical leadership and end-to-end ownership
- Ability to translate research into impactful systems
Skills
LLMs, Fine-Tuning, Alignment, Multi-Agent Systems, Tool Use, Retrieval, PyTorch, Rl, Long-Context Training, Text2Sql
Similar jobs
AI Research jobsLeads technical direction and develops maritime autonomy capabilities for unmanned surface and underwater vehicles, including motion planning, localization, safe behaviors, and heterogeneous multi-agent collaboration. Requires deep robotics and unmanned-systems experience, strong C++/Python skills, and senior technical leadership.
Leads design, implementation, integration, and field validation of tactical autonomy software for unmanned systems and multi-agent missions. Requires extensive autonomy or robotics experience, strong C++ and Python skills, and the ability to obtain a SECRET clearance.
Evaluates model and Generative AI risks across Upstart Bank’s model inventory, conducting risk assessments, monitoring reviews, quantitative analyses, and governance activities. Requires a quantitative master’s degree, 4+ years of relevant experience, and coding skills in Python, R, or similar languages.
Research and evaluate frontier AI capabilities for cybersecurity, rapidly prototyping tools, designing rigorous benchmarks, and helping operationalize reliable capabilities into products. Requires deep security expertise, strong technical communication, and at least seven years of relevant experience.
Own the architecture, delivery, evaluation, and production operations of AI capabilities embedded in procurement and finance workflows. The role requires 10+ years in applied AI or machine learning, deep LLM and agent expertise, and experience delivering measurable production outcomes.