Skip to content

Member of Technical Staff - Image / Video Generation

Trains and fine-tunes large-scale diffusion transformer models for image and video generation, conducts rigorous ablation studies, and optimizes distributed training. Requires hands-on diffusion-model experience, strong PyTorch and transformer expertise, and understanding of generative-model evaluation.

About the job

Responsibilities

  • Train large-scale diffusion transformer models for image and video data.
  • Rigorously ablate design choices by isolating variables, controlling for confounds, and communicating findings to shape research direction.
  • Evaluate speed-quality tradeoffs of neural network architectures in production settings.
  • Fine-tune diffusion models for specialized applications, including image and video upscalers, inpainting, and outpainting.
  • Debug distributed training issues and present research findings to the team.

Requirements

  • Hands-on experience training large-scale diffusion models for image and video data.
  • Practical knowledge of common failure modes and important training considerations.
  • Experience fine-tuning diffusion models for specialized applications.
  • Deep understanding of evaluating image and video generative models and selecting meaningful quality metrics.
  • Strong proficiency in PyTorch, transformer architectures, and modern deep learning.
  • Solid understanding of distributed training techniques, including FSDP, low-precision training, and model parallelism.

Nice-to-Haves

  • Experience writing forward and backward Triton kernels, ensuring correctness while considering floating-point errors.
  • Proficiency profiling, debugging, and optimizing single- and multi-GPU operations using tools such as Nsight and stack trace viewers.
  • Knowledge of the performance characteristics of architectural choices at scale.
  • Published research contributing to understanding of generative models.

Compensation

  • Base annual salary: €130,000–€340,000, plus equity.

Skills

Diffusion Models, PyTorch, Transformers, Deep Learning, Fsdp, Low-Precision Training, Model Parallelism, Triton, Nsight, Distributed Training, Image Generation, Video Generation

Protege

Protege

Remote

AI Engineer - New Verticals
No salary listedRemote3+ YOEML Engineering

Build the technical foundation for a new business vertical, creating reusable infrastructure and leading early customer engagements from scoping through delivery. The role requires 3+ years of engineering experience, strong Python and SQL skills, backend/data expertise, and comfort operating in ambiguity.

Adaption Labs

Adaption Labs

San Francisco, CA

Agent Systems Engineer
No salary listedHybrid5+ YOEML Engineering

Build production agent systems that plan, use tools, recover from failures, and improve over time. The role requires 5+ years of production ML or backend experience, LLM or agent deployment experience, and expertise in evaluation, tracing, observability, and agent architecture.

Adaption Labs

Adaption Labs

San Francisco, CA

Inference Performance Engineer
No salary listedHybrid5+ YOEML Engineering

Own inference-stack cost and performance by optimizing serving, caching, batching, quantization, decoding, routing, and GPU execution. The role requires 5+ years in ML systems, inference infrastructure, or performance engineering, plus strong Python and systems-language skills.

Black Forest Labs

Black Forest Labs

Freiburg, Germany

Member of Technical Staff - Post Training
€130k+/yrHybrid7+ YOEML Engineering

Owns end-to-end post-training for frontier multimodal generative models, spanning reward modeling, preference optimization, distillation, safety tuning, evaluation, and deployment. The role requires prior experience shipping post-training improvements and strong PyTorch expertise.

Datadog

Datadog

Bordeaux, France
Senior AI Engineer – Notebooks
No salary listedHybrid6+ YOEML Engineering

Build and operate AI-powered, customer-facing workflows for Datadog Notebooks, combining reliable backend systems with LLM capabilities. The role requires 6+ years of engineering experience, Go or Python expertise, and experience delivering production AI products.