Build and ship AI-native quality platform features, integrating and evaluating LLMs in real-world applications. The role requires strong software engineering fundamentals, 3+ years of experience, and hands-on expertise in prompt engineering, LLM observability, fine-tuning, and evaluation systems.
150k – 220k/yr
On-site3+ YOEML Engineering
About the role
Responsibilities
Solve difficult problems with product and technical ambiguity.
Build, ship, and take ownership of product functionality across the stack.
Integrate and apply large language models in real-world applications.
Work autonomously, manage project details, and deliver efficiently.
Requirements
Experience with LLM performance tuning, including prompt engineering and context management strategies.
Experience with LLM evaluations and observability.
Experience integrating LLMs into real-world applications using modern technologies such as Python and TypeScript.
Familiarity with supervised fine-tuning (SFT) and/or reinforcement learning fine-tuning (RLFT) on platforms such as Fireworks, Together AI, or Baseten.
Familiarity with evaluation dataset engineering, prompt engineering, LLM-as-a-judge verifiers, and reinforcement learning environments.
3+ years of experience.
Must be located in San Francisco or willing to relocate and work in person.
Nice-to-haves
Experience with supervised or reinforced fine-tuning.
Experience with classical machine learning techniques such as template matching, bounding box detection, and optical character recognition (OCR).
Experience running statistical experiments.
Technology Stack
React
TypeScript
Next.js
Node.js
PostgreSQL
Google Cloud
Kubernetes
Compensation and Benefits
Annual salary: $150,000–$220,000.
Competitive medical, vision, and dental insurance.
401(k).
Unlimited paid time off.
Company offsites and events.
Fully stocked kitchen.
Visa Sponsorship
Sponsorship available for TN, E-3, or J-1 visas.
H-1B visa and Green Card sponsorship or transfer are not currently available.
Build, deploy, and maintain machine and deep learning algorithms for biosignal-based medical devices, working across data curation, experimentation, validation, production, and client impact. Requires 4+ years of industry ML experience, DSP and statistics expertise, and proficiency with PyTorch or comparable frameworks.
150k – 170k/yrRemote4+ YOEML Engineering
Applied ML Engineer
DeepgramCalifornia
Own the research-to-production pipeline at Deepgram, turning experimental speech ML models into reliable, scalable production services. Partner with researchers on robust workflows, automated release gates, inference optimization, and feedback loops across hybrid GPU infrastructure.
150k – 220k/yrRemote5+ YOEML Engineering
Backend Engineer
DeepgramCalifornia
Backend Engineer building and optimizing Deepgram's core inference services for speech processing, including networking, audio transcoding, latency/memory optimization, and distributed compute orchestration. Requires 3+ years experience with Rust (or C/C++) and Python.
150k – 220k/yrRemote3+ YOEML Engineering
Perception & Fusion Engineer
Applied IntuitionArlington, VA +2
Develop and integrate real-time sensor perception and fusion algorithms for autonomous vehicles across land, air, sea, and space domains. Requires MS/PhD or 5+ years experience with multi-modal sensors (EO/IR/radar), ML deployment, and DoD customer collaboration.
150k – 220k/yrOn-site5+ YOEML Engineering
Research Engineer, Post-Training
Distyl AISan Francisco, CA +1
Research Engineers at Distyl build and productionize post-training techniques (fine-tuning, RLHF, reward models, evals) to improve reliability and behavior of compound AI systems for enterprise customers. Requires strong applied ML experimentation skills and ownership of real-world outcomes.