Skip to content
DominoDominoUnited States

Staff Software Engineer, MDLC

Staff Software Engineer building and scaling Domino's platform for multi-agent workflows, LLM inference infrastructure, and custom extensions. Requires deep backend/distributed systems experience plus familiarity with ML/AI workflows.

200k – 250k/yr
Remote7+ YOEML Engineering

About the role

What your impact will be

In your first year, you will:

  • Build and enhance platform features that enable teams to design, test, and deploy multi-agent workflows at scale.
  • Enhance Domino’s Extensions framework for enabling customers to build custom modules that extend platform feature and function.
  • Expand the platform's inference infrastructure to support high-throughput, low-latency serving of large language models, helping customers confidently operationalize LLM applications at enterprise scale.

What we look for in this role

  • Building Scalable Systems: Hands-on experience developing and managing high-performance back-end systems in distributed computing environments.
  • Collaboration Across Teams: Working closely with cross-functional teams to integrate systems with front-end interfaces and third-party services.
  • API Development: Designing and implementing secure, scalable APIs (e.g., RESTful APIs, gRPC).
  • Performance Optimization: Profiling and optimizing back-end performance, especially in cloud environments or with container technologies like Docker and Kubernetes.
  • Testing and CI/CD: Using robust testing frameworks (unit, integration, end-to-end) and setting up CI/CD pipelines.
  • Familiarity with traditional machine learning model development and AI workflows, including experiment tracking, hyperparameter optimization, model evaluation frameworks, and managing model artifacts.
  • Distributed Computing: Experience with frameworks like Apache Spark, Azure ML, or SageMaker is a plus.
  • Cloud Platforms: Proficiency with cloud providers (AWS, Azure, GCP) and deploying services in these environments.
  • Back-End Development: Expertise in languages such as Python, Java, Scala, or Go.

Skills

PythonJavaScalaGoKubernetesDockerAWSAzureGCPSparkREST APIsgRPCCI/CD

Similar roles

ML Engineering jobs
Phylo

Member of Technical Staff

PhyloSouth San Francisco, CA

Build and evaluate production AI agent harness systems for biomedical discovery. Own multi-agent coordination, planning, tool use, memory, rigorous evaluations, and infrastructure for reliable scientific agents. Requires strong backend/ML engineering and quantitative judgment.

200k – 300k/yr
On-site5+ YOEML Engineering
Clear Street

Senior / Staff AI Platform Engineer

Clear StreetUnited States

Build and own the core AI platform powering an AI-native trading copilot. Develop high-performance Rust backend for streaming, tool execution, and safe trading actions; design robust APIs with observability and security. Requires 8+ years systems programming experience.

200k – 350k/yr
Remote8+ YOEML Engineering
Clear Street

Senior / Staff AI Model Engineer

Clear StreetUnited States

Own reliability and quality for an AI copilot in a trading platform. Design evaluation systems, benchmarks, quality gates, model improvement loops, and AI monitoring for correctness, safety, and performance in market analysis and trading workflows. Requires 8+ years production software experience and strong ML eval expertise.

200k – 350k/yr
Remote8+ YOEML Engineering
Decagon

Staff Software Engineer, Agents

DecagonSan Francisco, CA

Build and own end-to-end AI agents for enterprise customers, integrating latest text/voice models and iterating based on real-world usage. Requires 8+ years of software engineering experience with Python and TypeScript.

200k – 400k/yr
On-site8+ YOEML Engineering
Nuance Labs

Member of Technical Staff - Research Fellow

Nuance LabsSeattle, WA

3-month research fellowship for early-career researchers working on frontier Multimodal LLMs, generative modeling, and real-time audiovisual AI. Own a research problem in pretraining, post-training, RL, evaluation, or multimodal modeling. Strong PyTorch and first-author tier-1 paper required.

200k – 250k/yr
On-siteML Engineering