# Member Of Technical Staff, Post-Training

**Company:** [Cohere](https://hotfix.jobs/companies/cohere)
**Location:** San Francisco, CA, New York, NY, London, United Kingdom, Paris, France, Toronto, Canada, Montreal, Canada
**Role:** ML Engineering
**Skills:** Python, JAX, PyTorch, Xla, Mlir, Kubernetes, Slurm, Ray, Distributed Training, Supervised Fine-Tuning, Reinforcement Learning, Model Post-Training, Performance Optimization, LLMs
**Posted:** 2025-06-13

> Build and post-train frontier AI models at scale, bridging research and production through scalable training software, distributed infrastructure, and performance optimization. The role requires strong software engineering skills and experience with large-model training and post-training.

## Job Description

## Responsibilities
- Design and write high-performance, scalable software for training models.
- Post-train models to reach state-of-the-art performance.
- Coordinate with specialist teams, including Agentic and Code, to produce models with strong overall performance.
- Develop and implement techniques to improve training-cycle performance across supervised fine-tuning (SFT) and reinforcement learning (RL).
- Research, implement, and experiment with ideas using large-scale compute and data infrastructure.
- Contribute to production code and research efforts.

## Requirements
- Extremely strong software engineering skills.
- Proficiency in Python and machine-learning frameworks including JAX, PyTorch, and XLA/MLIR.
- Experience with distributed training infrastructure such as Kubernetes and Slurm, and associated frameworks such as Ray.
- Experience using large-scale distributed training strategies.
- Hands-on experience training large models at scale.
- Hands-on experience with model post-training, with a strong emphasis on performance optimization.

## Nice-to-haves
- Publications at top-tier venues such as NeurIPS, ICML, ICLR, AIStats, MLSys, JMLR, AAAI, Nature, COLING, ACL, or EMNLP.

## Compensation and Benefits
- Weekly lunch stipend of $75/£75 or equivalent in local currency.
- Full health and dental benefits, including a separate mental-health budget.
- RRSP matching, 401(k), or pension scheme, depending on location.
- 100% parental-leave top-up for up to six months for either parent.
- Annual enrichment benefits covering arts and culture, fitness and wellness, quality time, and workspace improvements.
- Education and learning stipend for conferences, courses, and coaching.
- Six weeks of paid vacation (30 working days).
- Travel budget for remote employees visiting other offices and an annual company offsite.
- Coworking benefit for employees not near an office.
- $500 home-office setup stipend.

## Similar jobs

- [Performance Engineer, Inference Engine](https://hotfix.jobs/jobs/adc277be-252d-483e-9d49-cbc5716291f9) - Anthropic - San Francisco, CA - $350k – $850k/yr
- [Forward Deployed Engineer (AI and Automation)](https://hotfix.jobs/jobs/be838d7c-b8eb-4112-9f23-7f1b4665cb08) - OnePay - Remote - $150k – $190k/yr
- [Research Software Engineer, Post Training](https://hotfix.jobs/jobs/168e3c8f-8577-4482-bf94-91b3d11744ba) - Thinking Machines Lab - San Francisco, CA - $350k – $475k/yr
- [Software Engineer, AI for Chip Design](https://hotfix.jobs/jobs/abfb017d-1ce2-4d24-901d-16084eb7b3bc) - OpenAI - San Francisco, CA - $266k – $468k/yr
- [AI Software Engineer](https://hotfix.jobs/jobs/c9a0e889-a36e-488f-87d0-e9c27ab63bf0) - Rollstack - Remote

**Apply:** https://hotfix.jobs/jobs/8a0b139c-7f3b-4682-833f-1ec11ffb47b4
**Canonical:** https://hotfix.jobs/jobs/8a0b139c-7f3b-4682-833f-1ec11ffb47b4