Staff Software Engineer, AI Platform
Technical leader building agent infrastructure, observability, evals, and guardrails for production AI systems at Watershed. Requires 6+ years backend/platform/AI engineering experience and production TypeScript systems.
About the job
Responsibilities
- Design and build the agent infrastructure that powers Watershed's products
- Develop the observability and tracing layer for agent decisions, making it possible to debug, evaluate, and improve agent behavior at scale
- Build evals, harnesses, and guardrails that turn agent capabilities into production-grade, dependable systems
- Collaborate with product and other AI engineering teams to set product and technical strategy, and define the boundaries between autonomous agent behavior, deterministic code, and human oversight
- Keep up with developments and state-of-the-art in AI and agent infrastructure to determine what is relevant to Watershed
- Work closely with Watershed product teams to contribute your expertise to build agent experiences across the product
- Write performant, well-crafted, tested, and maintainable code across our technical stack
Requirements
- 6+ years of experience in backend, platform, or AI/ML engineering
- Experience building products and infrastructure that leverage LLMs, embeddings, and other ML technologies
- Full lifecycle experience building, deploying, and monitoring production systems that depend on LLMs or other ML technologies
- Experience with model evaluation, agent observability, and making non-deterministic systems reliable
- Experience building and operating production Typescript systems
- Must be willing to work from an office 4 days per week
Skills
TypeScript, LLMs, Embeddings, Machine Learning, Agent Infrastructure, Observability, Model Evaluation, Backend Engineering, Production Systems, Ai/Ml Engineering
Similar jobs
ML Engineering jobsDevelop production C++ perception capabilities for autonomous systems, spanning algorithms, libraries, integration, validation, and release. The role requires deep expertise in at least one perception domain, strong systems debugging, and experience delivering maintainable software in complex robotics or real-time environments.
Develop and productize online mapping models for autonomous navigation using real-world sensor data. The role requires deep ML expertise, robotics or computer vision experience, strong Python and deep learning framework skills, and a staff-level ability to deliver practical solutions.
Build and operate ML infrastructure for autonomy teams, including training and deployment pipelines, model observability, inference serving, and compiler platforms across hardware targets. Requires a degree, 3+ years of relevant experience, Python proficiency, and distributed-systems expertise.
Staff machine learning engineer leading scalable ranking, search, recommendation, and personalization systems. The role requires 9+ years of applied machine learning experience, strong programming and data engineering skills, and expertise productionizing models and pipelines.
Architects and operates production machine-learning systems that classify web and API traffic, detect bots and scrapers, and support real-time mitigation at internet edge latency. The role requires 9+ years of applied ML experience in adversarial domains and strong expertise in evaluation, data pipelines, and large-scale systems.