Machine Learning Engineer II, Computer Vision Applied Science
Build and fine-tune vision-centric VLMs and generative models using Pinterest's visual-text datasets. Requires 2+ years industry computer vision experience and an M.S. or Ph.D.
About the job
What you’ll do
- Prototype new model architectures for Pinterest VLMs. We’re looking for hands-on experience working with finetuning open-source LLM models and improve their visual perception and tool using capabilities.
- Develop new evaluation benchmarks that tailors to vision-centric capabilities such as fashion style recommendations.
- Read research papers, participate in group discussions, and help brainstorm our overall visual generative strategy at the company.
- Help with collection of relevant visual training data for Pinterest Canvas, particularly to conduct RLHF, targeted fine-tuning, etc.
- Publish and publicize your work via conferences, paper submissions, blog posts, etc.
- Mentor more junior researchers or research interns within the Pinterest Labs organization.
What we’re looking for
- Research engineers and scientists who have experience working with generative computer vision models, preferably various forms of visual encoders and LLMs.
- 2+ years of industry computer vision experience.
- M.S. or PhD in Machine Learning, Computer Science, or related areas.
Nice to Have
- Publications at top ML conferences.
- Experience using Cursor, Copilot, Codex, or similar AI coding assistants for development, debugging, testing, and refactoring.
- Familiarity with LLM-powered productivity tools for documentation search, experiment analysis, SQL/data exploration, and engineering workflow acceleration.
Skills
Computer Vision, Machine Learning, LLMs, Generative Models, Visual Encoders, Finetuning, RLHF, Multimodal Models, PyTorch, TensorFlow
Similar jobs
ML Engineering jobsDevelop and deploy responsible AI and machine learning fairness solutions across Pinterest’s large-scale, user-facing products, including generative AI, search, and recommendations. The role requires production ML experience, expertise in fairness interventions and modern architectures, and a master’s or PhD in computer science or a related field.
Build production infrastructure for replayable enterprise environments, agent evaluation, and continuous model improvement. The role combines hands-on customer deployment, research experimentation, large-scale data processing, and production software engineering.
Design, deploy, and improve real-time machine learning systems for Lyft’s ride fulfillment and marketplace products. The role requires 2+ years of ML experience, production programming skills, and expertise with deep learning and recommendation systems.
Build datasets, evaluations, and scalable data systems that improve frontier AI models on challenging biological and scientific tasks. The role partners with scientists and AI labs and requires at least two years of experience applying biology and AI, plus hands-on LLM experience.
Builds high-scale data pipelines, distributed systems, and AI agent workflows using LLMs for fraud intelligence platform. Requires 2+ years software engineering, Python proficiency, big data tools, AWS/K8s, and ML foundations.