Skip to content
StreamStream

Staff Backend Engineer – AI

Own end-to-end development, evaluation, and production deployment of AI models serving high-volume real-time products. The role requires 5+ years of production Python experience, hands-on fine-tuning and ML operations, cloud infrastructure expertise, and strong technical ownership.

About the job

Responsibilities

  • Own development, fine-tuning, and evaluation of in-house AI models from dataset design through production deployment.
  • Run supervised fine-tuning and post-training experiments; establish benchmarks and evaluation harnesses.
  • Build and maintain training and evaluation data pipelines, including data quality, labeling, and reproducibility.
  • Deploy models on Stream’s serving stack and optimize latency, cost, and reliability at high volume.
  • Set technical direction for ambiguous AI problems and decide what to build, test, or abandon.
  • Integrate models with Go-based API teams and infrastructure across the engineering organization.
  • Contribute to open-source projects and share work through code, writing, or community engagement.
  • Raise engineering standards through code review, mentorship, and pragmatic best practices.

Requirements

  • 5+ years of production-level Python engineering experience.
  • Hands-on experience with supervised fine-tuning and post-training of models.
  • Familiarity with fine-tuning and serving tools such as Unsloth, Fireworks, Baseten, or equivalent technologies.
  • Experience with GCP or AWS, including infrastructure as code with Terraform.
  • Experience operating ML-based products in production, including deployment, monitoring, retraining, and iteration.
  • Experience designing and operating data pipelines for training and evaluation.
  • Demonstrated ownership of ambiguous problems from definition through delivery.
  • Strong communication skills and comfort working in a small, distributed, fast-moving team.

Nice to Have

  • Visible open-source contributions or maintained libraries.
  • Experience with Go.
  • Deep understanding of Python concurrency and asyncio limitations in high-throughput systems.
  • Experience with real-time or low-latency inference systems.
  • Experience as an early engineer, founder, or in a startup environment with an undefined roadmap.

Compensation & Benefits

  • 28 days paid time off plus paid Dutch holidays.
  • Company equity.
  • Pension scheme.
  • Learning and Development budget.
  • Commute expenses to Amsterdam covered or company bike option within the city.
  • Fitness stipend.
  • MacBook Pro.
  • Healthy team lunches and snacks.
  • Generous relocation package.
  • Benefits are adjusted according to the employee’s location of residence.

Skills

Python, Machine Learning, Supervised Fine-Tuning, Terraform, GCP, AWS, Go, Unsloth, Fireworks, Baseten, Data Pipelines, Asyncio, Model Serving, Open Source

Payabli

Payabli

Remote

Staff Machine Learning Engineer
No salary listedRemote8+ YOEML Engineering

Sets the technical direction for production machine learning across a payments platform, building and scaling models for risk, authorization, disputes, and forecasting. Requires 8+ years of ML engineering experience, including production model ownership and strong technical leadership.

1Password

1Password

United States
Staff AI Marketing Systems Engineer
No salary listedRemote7+ YOEML Engineering

Build and operate production AI agent systems that help Sales and Marketing teams with account planning, deal support, competitive intelligence, and content creation. The role requires 6+ years of experience shipping reliable LLM workflows with retrieval, tool use, permissions, evaluation, and observability.

Okta

Okta

Toronto, Canada

Staff Machine Learning Engineer, Generative AI
CA$168k+/yrHybrid7+ YOEML Engineering

The Staff Machine Learning Engineer will architect and deploy scalable generative AI and machine learning systems, including retrieval, inference, evaluation, and agentic workflows. The role requires 7+ years of software development experience, strong Python skills, applied ML expertise, and deep familiarity with modern GenAI platforms and frameworks.

Grafana Labs

Grafana Labs

United States
Staff AI Engineer
CA$164k+/yrRemote8+ YOEML Engineering

Builds and owns production multi-agent AI infrastructure, backend integrations, and workflow automation for marketing operations. Requires 8+ years of software engineering experience, strong Python and JavaScript/Node.js skills, production LLM experience, and deep Google Cloud expertise.

Dialpad

Dialpad

Vancouver, Canada

Senior AI Engineer
CA$185k+/yrOn-site5+ YOEML Engineering

Senior AI engineer owning production systems for real-time speech models and AI voice agents. The role requires 5+ years of production software experience, Python, model serving and inference optimization, cloud and distributed systems expertise, and strong reliability and operations skills.