Machine Learning Engineer
Build and operate machine learning models for sales roleplay, scoring, and coaching products, owning the lifecycle from fine-tuning and evaluation through production and on-device deployment. The role emphasizes open-source models, latency and privacy optimization, and rigorous model testing.
About the job
Responsibilities
- Own machine learning models end to end, from training and fine-tuning through production deployment and ongoing operation.
- Fine-tune and deploy open-source models to optimize cost, latency, and capabilities.
- Run models on-device when customer latency or privacy requirements demand it.
- Apply model quantization and distillation tradeoffs for on-device deployment.
- Build evaluation frameworks, benchmarks, and regression suites to measure model changes.
- Develop and ship models supporting roleplay, scoring, and coaching products based on real sales calls.
- Collaborate with founders and engineering on product direction.
Compensation and Benefits
- Compensation: $260,000–$300,000+ based on experience, plus meaningful equity.
- Medical, dental, and vision insurance.
- 401(k).
- Commuter and parking benefits.
- Unlimited paid time off.
- Free lunch and dinner in the office.
Skills
Machine Learning, Open-Source Models, Model Fine-Tuning, Model Deployment, Model Quantization, Model Distillation, On-Device Machine Learning, Evaluation Frameworks, Benchmarking, Regression Testing
Similar jobs
ML Engineering jobsBuild research infrastructure and tooling that enables AI models to design silicon, including reinforcement learning environments, EDA integrations, evaluations, and experiment workflows. The role requires strong software engineering fundamentals and comfort working across research, tooling, and chip-design systems.
Build and optimize the production LLM inference runtime for frontier models on OpenAI’s custom silicon. The role spans scheduling, distributed execution, memory and KV-cache management, performance tooling, and hardware-software co-design.
Build and deploy LLM-powered tools, agents, and ecosystem infrastructure with life sciences research institutions. The role requires deep scientific or biomedical research experience, production software development expertise, and the ability to translate partner workflows into scalable AI systems.
Build and deploy algorithmic systems for high-impact healthcare problems, choosing among machine learning, optimization, heuristics, and hybrid approaches. The role requires 4+ years of relevant industry experience, strong applied problem-solving and evaluation skills, and fluency in modern ML tooling.
Build and optimize OpenAI’s inference stack for AWS Trainium across high-performance kernels, compilers, runtimes, and model execution. The role requires systems programming and accelerator experience, with opportunities to solve end-to-end performance problems for frontier-scale AI models.