Applied AI, Research Engineer
Applied AI Research Engineer who tests model capabilities, builds demos and evaluations, supports strategic customer implementations, and translates field insights into product and research direction. Requires 6+ years of technical experience, programming proficiency, LLM development experience, and strong communication skills.
About the job
Responsibilities
- Embed with product teams to become the field’s technical expert on new products and capabilities.
- Own the technical go-to-market handshake for model and product launches, including early testing, field-readiness assessments, and capability education.
- Build demos, evaluations, and prototypes that showcase customer-relevant capabilities and define performance standards for important domains.
- Create playbooks, training sessions, and reference architectures with Technical Enablement.
- Solve novel technical challenges for strategic customers and package learnings into scalable approaches.
- Synthesize customer insights, adoption patterns, blockers, and capability gaps to inform product and research roadmaps.
- Partner with Sales and Applied AI teams on high-value accounts requiring specialized domain knowledge.
- Occasionally travel to customer sites for workshops, implementation support, and research collaboration.
Requirements
- 6+ years of experience as a software engineer, technical product manager, forward-deployed engineer, AI startup founder, or similar technical professional.
- Strong proficiency in at least one programming language.
- Meaningful experience building with large language models, including LLM-powered products, agents, workflows, or production integrations.
- Experience designing evaluations, building prototypes, or developing technical enablement for complex products.
- Ability to develop domain expertise quickly and become a credible technical voice.
- Excellent communication and interpersonal skills, including the ability to explain complex topics to diverse internal and external stakeholders and executives.
- Ability to execute amid ambiguity and adapt across domains.
- Bachelor’s degree or equivalent combination of education, training, and experience.
Nice-to-haves
- Customer or end-user exposure.
- Experience translating technical discoveries into materials, training, or tools that help others succeed.
- Passion for safe and beneficial applications of AI.
Compensation
- Annual salary: $300,000–$400,000 USD.
Skills
Python, LLMs, AI Agents, AI Workflows, Llm Integration, Model Evaluation, Prototyping, Technical Enablement, Reference Architectures, Product Development
Similar jobs
AI Research jobsLeads the research agenda for humanoid robotics, developing foundation-model and reinforcement-learning methods for dexterous manipulation and deploying them on real robotic systems. Requires a PhD, strong robotics research publications, and senior-level technical leadership.
Research and build safety models, evaluations, and runtime safeguards for conversational AI agents, addressing prompt injection, unsafe tool use, privacy, and policy risks. Requires 4+ years in AI/ML engineering, research, or safety plus experience deploying and evaluating language models or agentic systems.
Build and expand customer-facing agentic AI products, MCP integrations, and automated reconciliation workflows for private fund management. The role requires senior-level software engineering, strong systems thinking, product judgment, and hands-on experience building and evaluating AI systems.
Build Vanta’s organizational intelligence layer by shipping prototypes, internal tools, and AI agent workflows that make cross-source data useful to EPD, GTM, and other teams. The role requires recent hands-on LLM product work, independent problem scoping, and strong judgment around AI quality, reliability, cost, and latency.
Conduct research and build open foundation models and training systems aimed at accelerating scientific discovery. The role requires a PhD-level background and substantial experience training foundation models, with expertise in agentic training or multimodal data preferred.