Member of Technical Staff, Machine Learning
Machine Learning Engineer owning the full ML lifecycle for multimodal video datasets at Sieve. Fine-tune VLMs, build evaluation/QA pipelines with frontier models, design filtering systems over internet-scale data, and ship production improvements for top AI labs. Requires strong Python, PyTorch, and production ML experience.
About the job
What You'll Do
- Own model quality for customer-facing video understanding problems
- Fine-tune vision-language and multimodal foundation models for specialized tasks
- Build automated evaluation and QA pipelines using frontier models like Gemini, GPT, Claude, and open-source VLMs
- Design high-precision filtering, ranking, retrieval, and labeling systems over internet-scale video datasets
- Create datasets, benchmarks, and evaluation frameworks that continuously improve model quality
- Develop production ML pipelines spanning preprocessing, inference, post-processing, and quality validation
- Work directly with frontier AI labs to translate ambiguous requirements into scalable ML systems
- Ship improvements quickly, measure results, and iterate based on real-world performance
Requirements
- Strong Python engineer with experience building production ML systems
- Experience training, fine-tuning, or deploying modern deep learning models
- Comfortable working with PyTorch and modern foundation models
- Excellent intuition for evaluation, dataset quality, precision/recall tradeoffs, and edge cases
- Enjoys rapidly prototyping with new AI models and APIs
- Comfortable owning projects from customer problem to internal pipelines to deployed solution
- Strong communicator who enjoys working directly with customers and cross-functional teams
- Excited by video, multimodal AI, and frontier foundation models
Nice-to-Haves
- In-person at our SF HQ (all roles require onsite in San Francisco 5 days per week)
Skills
Python, PyTorch, Deep Learning, Multimodal Models, Vision-Language Models, Fine-Tuning, Model Evaluation, Gemini, Gpt, Claude, Vlm, Ml Pipelines
Similar jobs
ML Engineering jobsBuild reinforcement-learning environments, evaluations, datasets, and scalable infrastructure for frontier AI capabilities. The role suits a high-agency generalist engineer with experience in agents, evaluations, or RL workflows and strong communication skills.
Build, deploy, and optimize AI applications and machine learning models for customer use cases while contributing to an internal ML platform. This new graduate role requires a technical master’s degree, hands-on ML or LLM experience, and strong customer communication skills.
Build production infrastructure for replayable enterprise environments, agent evaluation, and continuous model improvement. The role combines hands-on customer deployment, research experimentation, large-scale data processing, and production software engineering.
Builds and owns production multi-agent AI infrastructure, backend integrations, and workflow automation for marketing operations. Requires 8+ years of software engineering experience, strong Python and JavaScript/Node.js skills, production LLM experience, and deep Google Cloud expertise.
Build replayable enterprise environments, evaluation systems, graders, and post-training workflows for AI agents. The role spans machine-learning research and production engineering and requires 1–7 years of software or ML systems experience.