Staff Applied AI Engineer, AI Experiences & Agents
Build and ship production-grade AI agents and experiences that help creators personalize, automate, and grow. Requires strong product engineering skills, hands-on LLM/agent experience, and rigorous evaluation practices.
About the job
What You’ll Own
- Build and ship creator-facing and visitor-facing AI experiences end-to-end, from prototype through production.
- Design multi-agent workflows that combine prompting, tool use, retrieval, memory/state, orchestration, and human-in-the-loop controls where appropriate.
- Create eval suites, regression harnesses, and quality metrics so AI quality improves deliberately rather than by anecdote.
- Partner closely with product, design, data, and trust teams to translate ambiguous user problems into simple, high-value experiences.
- Optimize latency, reliability, cost, and safety so AI features are fast, dependable, and viable at scale.
- Turn production feedback, support signals, and behavioral data into model, prompt, policy, and product improvements.
Who We’re Looking For
- Strong software engineering fundamentals and a track record of shipping production systems end-to-end.
- Hands-on experience building LLM-powered products, agents, chat systems, RAG workflows, or similar AI experiences.
- Experience building evals for AI output quality harnessing.
- Strong product sense: you can identify the real user problem, simplify the solution, and make practical trade-offs around effort, scope, and quality.
- Ability to frequently weigh trade-offs and understand the benefits of keeping systems architecture simple.
- Good judgment on safety, trust, privacy, abuse, and failure modes in AI systems.
- Strong communicator who collaborates well with cross-functional partners and explains trade-offs clearly.
Skills
LLMs, AI Agents, RAG, Prompt Engineering, Multi-Agent Systems, Evaluation Frameworks, Python, Production Systems, Latency Optimization, Safety & Trust
Similar jobs
ML Engineering jobsLeads the roadmap and technical vision for Snowflake Feature Store, building reliable, high-performance machine learning platform capabilities and supporting technical execution across partner teams. Requires 10+ years of experience with data-serving infrastructure or ML platforms, plus Java and Python expertise.
Leads the design, implementation, integration, and field validation of tactical autonomy and multi-agent coordination capabilities for unmanned platforms. Requires 7+ years of relevant experience, production C++, technical leadership, and eligibility for a U.S. Secret clearance.
Senior Staff ML Engineer fine-tunes and optimizes state-of-the-art LLMs for Airbnb's customer support AI products, including AI assistants and autonomous agents. Partners cross-functionally to productionize models at scale. Requires PhD and 10+ years experience with PyTorch.
Leads the design and operation of reliable, scalable model infrastructure powering AI inference across multiple providers. Requires 7+ years of distributed-systems engineering experience, strong programming skills, and expertise in production reliability and cloud infrastructure.
Staff Machine Learning Engineer building and operating production ML systems for causal marketing measurement, optimization, and planning. The role requires deep statistical and machine learning expertise, production programming experience, cross-functional collaboration, and technical mentorship.