Staff Research Engineer, Data Agents
Develop post-training recipes and build enterprise data agents for autonomous planning, code generation, and multi-step workflows. Requires 2+ years applied research experience shipping prototypes, plus expertise in LLMs, agents, and RL.
About the job
Responsibilities
- Develop the best post-training recipes to train enterprise Data agents, that are capable of autonomous planning, code generation, and multi-step workflow execution within intricate enterprise settings.
- Partner closely with product teams to turn prototypes and research ideas into the best agentic experience for Databricks users.
- Build systems that help the agent discover and use relevant lakehouse context, including tables, notebooks, code, and cell outputs, to produce more accurate and useful results.
- Raise the technical bar for the team through strong design, execution, debugging, and mentorship, helping shape the long-term direction of agentic experiences at Databricks.
Requirements
- BS, MS, or PhD in Computer Science or a related field.
- 2+ years of experience in an applied research environment, with a track record of shipping research prototypes to production.
- Experience in LLMs, agents, reinforcement learning, post-training workflows.
- Ability to work effectively in a fast-moving environment that blends research exploration with product and engineering rigor.
- Clear communication and strong cross-functional collaboration with researchers, engineers, and product stakeholders.
Skills
LLMs, Reinforcement Learning, Post-Training, Python, Machine Learning, AI Agents, Code Generation, Multi-Agent Systems
Similar jobs
AI Research jobsLeads technical direction and develops maritime autonomy capabilities for unmanned surface and underwater vehicles, including motion planning, localization, safe behaviors, and heterogeneous multi-agent collaboration. Requires deep robotics and unmanned-systems experience, strong C++/Python skills, and senior technical leadership.
Evaluates model and Generative AI risks across Upstart Bank’s model inventory, conducting risk assessments, monitoring reviews, quantitative analyses, and governance activities. Requires a quantitative master’s degree, 4+ years of relevant experience, and coding skills in Python, R, or similar languages.
Leads design, implementation, integration, and field validation of tactical autonomy software for unmanned systems and multi-agent missions. Requires extensive autonomy or robotics experience, strong C++ and Python skills, and the ability to obtain a SECRET clearance.
Research and evaluate frontier AI capabilities for cybersecurity, rapidly prototyping tools, designing rigorous benchmarks, and helping operationalize reliable capabilities into products. Requires deep security expertise, strong technical communication, and at least seven years of relevant experience.
Own the architecture, delivery, evaluation, and production operations of AI capabilities embedded in procurement and finance workflows. The role requires 10+ years in applied AI or machine learning, deep LLM and agent expertise, and experience delivering measurable production outcomes.