Skip to content
SieveSieve

Member of Technical Staff, Machine Learning

Machine Learning Engineer owning the full ML lifecycle for multimodal video datasets at Sieve. Fine-tune VLMs, build evaluation/QA pipelines with frontier models, design filtering systems over internet-scale data, and ship production improvements for top AI labs. Requires strong Python, PyTorch, and production ML experience.

About the job

What You'll Do

  • Own model quality for customer-facing video understanding problems
  • Fine-tune vision-language and multimodal foundation models for specialized tasks
  • Build automated evaluation and QA pipelines using frontier models like Gemini, GPT, Claude, and open-source VLMs
  • Design high-precision filtering, ranking, retrieval, and labeling systems over internet-scale video datasets
  • Create datasets, benchmarks, and evaluation frameworks that continuously improve model quality
  • Develop production ML pipelines spanning preprocessing, inference, post-processing, and quality validation
  • Work directly with frontier AI labs to translate ambiguous requirements into scalable ML systems
  • Ship improvements quickly, measure results, and iterate based on real-world performance

Requirements

  • Strong Python engineer with experience building production ML systems
  • Experience training, fine-tuning, or deploying modern deep learning models
  • Comfortable working with PyTorch and modern foundation models
  • Excellent intuition for evaluation, dataset quality, precision/recall tradeoffs, and edge cases
  • Enjoys rapidly prototyping with new AI models and APIs
  • Comfortable owning projects from customer problem to internal pipelines to deployed solution
  • Strong communicator who enjoys working directly with customers and cross-functional teams
  • Excited by video, multimodal AI, and frontier foundation models

Nice-to-Haves

  • In-person at our SF HQ (all roles require onsite in San Francisco 5 days per week)

Skills

Python, PyTorch, Deep Learning, Multimodal Models, Vision-Language Models, Fine-Tuning, Model Evaluation, Gemini, Gpt, Claude, Vlm, Ml Pipelines

Roboflow

Roboflow

San Francisco, CA

Member of Technical Staff — Frontier Data
$150k+/yrRemoteML Engineering

Build reinforcement-learning environments, evaluations, datasets, and scalable infrastructure for frontier AI capabilities. The role suits a high-agency generalist engineer with experience in agents, evaluations, or RL workflows and strong communication skills.

Fireworks AI

Fireworks AI

San Mateo, CA
Member of Technical Staff
$160k+/yrOn-siteEntry levelML Engineering

Build, deploy, and optimize AI applications and machine learning models for customer use cases while contributing to an internal ML platform. This new graduate role requires a technical master’s degree, hands-on ML or LLM experience, and strong customer communication skills.

Ambral

Ambral

New York, NY
Member of Technical Staff
$140k+/yrOn-siteEntry levelML Engineering

Build production infrastructure for replayable enterprise environments, agent evaluation, and continuous model improvement. The role combines hands-on customer deployment, research experimentation, large-scale data processing, and production software engineering.

Grafana Labs

Grafana Labs

United States
Staff AI Engineer
CA$164k+/yrRemote8+ YOEML Engineering

Builds and owns production multi-agent AI infrastructure, backend integrations, and workflow automation for marketing operations. Requires 8+ years of software engineering experience, strong Python and JavaScript/Node.js skills, production LLM experience, and deep Google Cloud expertise.

Ambral

Ambral

New York, NY
Member of Technical Staff
$165k+/yrOn-site1+ YOEML Engineering

Build replayable enterprise environments, evaluation systems, graders, and post-training workflows for AI agents. The role spans machine-learning research and production engineering and requires 1–7 years of software or ML systems experience.