Senior Applied AI Engineer
Senior Applied AI Engineer building the core intelligence layer for Roger, an AI platform for home health clinicians. Responsibilities include training/fine-tuning LLMs on proprietary clinical data, building rigorous eval and monitoring systems, and shipping reliable agentic LLM workflows that improve patient care.
About the job
What You'll Do
- Train and fine-tune open source models, leveraging our vast proprietary dataset to push accuracy beyond what off-the-shelf models can do.
- Build eval datasets and pipelines that let us measure model accuracy rigorously and improve it continuously.
- Design how we measure accuracy in the first place: the metrics, harnesses, and feedback loops that turn real clinical outcomes into measurable model improvements.
- Build scalable, cost-efficient inference infrastructure with great monitoring and observability.
- Build better agentic infrastructure and partner on the interfaces that turn model capability into a great clinician experience.
- Stay at the frontier: keep up with the latest research, frontier model capabilities, and open source frameworks, and bring the best of it into production.
- Prototype quickly, then harden into scalable, secure, and reliable production systems.
You Might Be a Good Fit If You Have
- 7+ years of professional software engineering experience, with meaningful depth in AI/ML.
- Experience training, fine-tuning, or evaluating LLMs and open source models, with real opinions about what works and what does not.
- Shipped real AI software to real users. You can describe something you built, what broke, and how you fixed it.
- Strong instincts for evals, observability, and the feedback loops that turn user feedback into measurable improvement.
- Experience building agentic systems: tool use, generator and critic loops, planners and executors, and orchestration where one agent's output drives another's work.
- Comfort building infrastructure that is fast, reliable, and cost-efficient at scale.
- Startup experience shipping real features in high-growth environments.
- A product mindset, comfort across the stack, and the ability to operate in ambiguity without a clean spec.
- High standards for reliability and accuracy when real clinicians and patients depend on your work.
Why Join Us?
- Platinum health, dental, and vision insurance
- Flexible PTO
- Be at the forefront of LLM innovation in healthcare
- Founded by AI Researchers from Cornell and Amazon Alexa
- Make a tangible difference in the lives of clinicians and patients
- Work on a mission that matters, backed by AI innovation and top-tier investors
Skills
LLMs, Fine-Tuning, Model Evaluation, Agentic Systems, Inference Infrastructure, Observability, Python, PyTorch, Transformers
Similar jobs
ML Engineering jobsBuild and deploy production AI-agent systems, including their harnesses, evaluations, orchestration, and supporting services. The role requires 5+ years of software engineering experience, production LLM or agent experience, and strong Python or TypeScript/Node.js skills.
Own machine learning end to end, from modeling messy clinical data through production deployment, monitoring, and infrastructure. The role requires 7+ years of experience building scalable ML systems, strong software and data engineering skills, and proficiency with Python, SQL, and cloud platforms.
Leads and manages an applied machine learning team developing production fraud detection and identity verification models. The role combines people leadership with hands-on technical work and requires substantial ML experience, production deployment expertise, and experience in risk-focused domains.
Builds and productionizes machine learning systems for trust and safety, including abuse detection, autonomous AI agents, and evaluation frameworks. The role requires 5+ years of applied ML experience, strong Python skills, experience with LLMs and scalable pipelines, and a relevant advanced degree or equivalent background.
Owns end-to-end production machine learning systems, including NLP, LLM, agentic, ranking, and recommendation capabilities. Requires 8+ years of industry experience, strong Python and cloud ML expertise, and the ability to deliver explainable AI products with cross-functional and customer impact.