Skip to content
StackblitzStackblitz

Senior Applied AI Engineer

Designs and implements AI agents that transform natural language into production-ready full-stack applications using state-of-the-art LLMs. Integrates multiple LLM providers, orchestrates workflows, and continuously improves agent performance through data analysis and experimentation. Requires TypeScript proficiency and hands-on LLM experience.

About the job

How You'll Contribute

  • Develop AI Agents: Design and implement AI agent features and extend existing agents with new capabilities. This includes managing the agent’s context (using techniques like sub-agents, retrieval based context management, sliding context windows, etc.) so it can handle long conversations or large knowledge and code bases efficiently.
  • Integrate Multiple LLM Providers: Leverage models from providers such as OpenAI (GPT series), Anthropic (Claude), and Google (Gemini). Quantitatively evaluate and choose the best model for a given task, and incorporate new model features or improvements (often by beta-testing new releases and assessing their strengths).
  • Tool Use and Workflow Orchestration: Enable the AI agent to call external tools and APIs safely and effectively. Implement structured approaches to allow the agent to perform actions like web searches, database queries, fetch additional information, or other domain-specific operations. Utilize frameworks such as Vercel’s AI SDK, LangGraph and others for building multi-step AI workflows.
  • Team Collaboration: Work closely with your immediate team and adjacent teams to deliver AI-powered features. Collaborate effectively with peer engineers and product managers to ensure AI-driven features are production-ready, efficient, maintainable, and well-monitored in deployment. Share knowledge and help lift up mid-level and junior engineers on the team.
  • Data Collection and Analysis: Collect and curate datasets from agent responses and multi-turn conversations to understand agent behavior. Analyze conversation patterns, failure modes, and success signals to derive actionable insights that drive improvements to agent performance and user experience.
  • Continuous Improvement and Evaluation: Stay up-to-date with the latest research in NLP and LLMs, and experiment with novel techniques (e.g. new prompting strategies, context handling methods, model fine-tuning opportunities). Continuously evaluate the AI system’s performance using systematic tests and user feedback, and iterate on prompts, agents and workflows to improve output quality and reliability (for example, by developing automated LLM evaluation benchmarks).

Qualifications

  • TypeScript: Familiarity with TypeScript is important. Our entire stack is built on it. Willingness to work in TS daily is key.
  • LLM Experience: Hands-on experience working with Large Language Models (LLMs) and understanding their capabilities and limitations. Proven experience building applications or systems powered by LLMs.
  • Prompt Engineering: Deep understanding of prompt engineering best practices to guide LLM behavior. Able to craft, refine, and optimize prompts for different tasks and models.
  • Software Engineering Skills: Solid software engineering fundamentals with experience in building production-ready systems.
  • Autonomous Execution: Ability to deliver high-quality work with minimal oversight. Comfortable owning tasks end-to-end, managing complexity, and driving projects to completion within your functional area.
  • Problem-Solving: Strong analytical and problem-solving skills with the ability to debug complex AI behaviors and bring clarity to ambiguous tasks.
  • Data-Driven Mindset: Comfortable collecting, curating, and analyzing data to inform decisions. Able to build datasets, identify patterns in agent behavior, and translate findings into actionable improvements.
  • Strong verbal and written English communication skills are required.

Bonus Points

  • DSPy Framework: Familiarity with DSPy (Declarative Self-improving Python) for building modular AI systems and optimizing prompts programmatically.
  • Machine Learning Background: Understanding of ML fundamentals and experience with model evaluation metrics.
  • Open Source Contributions: Experience contributing to or maintaining open-source AI/ML projects.
  • Research Background: Experience reading and implementing techniques from AI/ML research papers.

Skills

TypeScript, LLMs, Prompt Engineering, OpenAI, Anthropic Claude, Google Gemini, Vercel Ai Sdk, LangGraph, Dspy, Machine Learning

Mercury

Mercury

San Francisco, CA
Senior Machine Learning Operations Engineer
$157k+/yrRemote5+ YOEML Engineering

Build and operate the platform that deploys, serves, observes, and retrains production machine-learning models for real-time fraud and financial-crime risk decisions. Requires 5+ years of ML engineering, backend, or MLOps experience, strong Python skills, and production model-serving expertise.

Traba

Traba

New York, NY
Senior Software Engineer
$200k+/yrOn-site5+ YOEML Engineering

Build and deploy production AI-agent systems, including their harnesses, evaluations, orchestration, and supporting services. The role requires 5+ years of software engineering experience, production LLM or agent experience, and strong Python or TypeScript/Node.js skills.

Front

Front

San Francisco, CA

Senior Applied AI Engineer
$205k+/yrHybrid5+ YOEML Engineering

Build and deploy generative AI and LLM-powered agentic applications at Front to automate customer support inquiries, enhance product capabilities, and drive operational insights. Requires 5+ years software engineering experience with strong production AI/ML focus, agentic/RAG expertise, and proficiency in Node.js, TS, and Python.

Baselayer

Baselayer

San Francisco, CA

Senior AI Engineer, Agentic Data Enrichment
$230k+/yrHybrid5+ YOEML Engineering

Senior AI Engineer responsible for production LLM agents that enrich business identity data through web discovery, verification, classification, and risk scoring. The role requires strong asynchronous Python, agent and evaluation expertise, browser automation, and experience operating AI systems in production.

Blee

Blee

San Francisco, CA

Senior AI Engineer
$150k+/yrHybrid5+ YOEML Engineering

Designs and ships production multi-agent compliance systems, including LLM pipelines, model training, evaluation, monitoring, and explainability. Requires 5+ years of applied AI/ML engineering experience, strong Python, and experience deploying production ML systems.