Skip to content
PostmanPostman

Applied AI Scientist, Small Language Model and AI Training

Leads R&D on small language models and AI training, developing efficient architectures, optimizing performance, and ensuring safety. Collaborates with research, engineering, and product teams using Python, PyTorch, TensorFlow, or JAX.

About the job

The Opportunity

As an Applied Scientist specializing in Small Language Models and AI Training, you will lead research and development efforts focused on building efficient, high-performance language models tailored for practical applications. You will work closely with research, engineering, and product teams to advance model training techniques, optimize architectures, and scale AI solutions. Your work will directly contribute to AI systems that are safe, interpretable, and impactful across diverse usage scenarios.

What You’ll Do

  • Lead research and development of novel training methodologies and architectures for small and efficient language models.
  • Design, implement, and evaluate model training experiments to improve performance, robustness, and generalization of language models.
  • Collaborate closely with research scientists and engineers on scalable training pipelines and model deployment strategies.
  • Develop techniques for model compression, fine-tuning, and domain adaptation to optimize models for real-world applications.
  • Ensure AI safety, fairness, and alignment principles are integrated into model training processes and evaluated rigorously.
  • Mentor and support cross-functional teams on applied machine learning methods and best practices.
  • Evaluate and integrate new tools, frameworks, and datasets to accelerate AI training workflows.
  • Partner with product teams to translate model capabilities into actionable features aligned with user needs and ethical standards.

About You

  • Have demonstrated experience in applied research or engineering roles focused on training language models, ideally small or efficient models.
  • Strong programming skills in Python and familiarity with machine learning frameworks such as PyTorch, TensorFlow, or JAX.
  • Deep understanding of language model architectures, training techniques, and optimization strategies.
  • Experience with distributed training, data pipeline design, and scalable AI infrastructure.
  • Passion for AI safety, interpretability, and delivering user-centered AI technology.
  • Excellent communication skills with proven ability to collaborate across research, engineering, and product teams.

Preferred

  • Prior experience working with large and small language models in production or research settings.
  • Background in reinforcement learning, prompt engineering, or transfer learning techniques.
  • Experience with developer tools, APIs, or frameworks related to AI model integration and delivery.
  • Knowledge of AI alignment, fairness, and ethical AI training methodologies.

Compensation: Base salary $218,500 - $276,000 plus equity.

Skills

Python, PyTorch, TensorFlow, JAX, Language Models, Distributed Training, Model Compression, Fine-Tuning, Reinforcement Learning, Ai Safety

Baseten

Baseten

San Francisco, CA

AI Engineer
$220k+/yrHybrid5+ YOEAI Research

Build and ship agentic AI product experiences, internal automation, and customer-facing features across the stack. The role requires 5+ years of software engineering experience, hands-on experience with AI or LLM-powered products, Python proficiency, and strong autonomy.

Mercor

Mercor

San Francisco, CA

Research Scientist, APEX Benchmarks
$200k+/yrOn-siteAI Research

Leads the design, measurement, publication, and adoption of APEX benchmarks evaluating frontier models on economically valuable professional work. The role requires rigorous research judgment, strong coding and statistical skills, and excellent communication across technical, commercial, and research audiences.

Tessera Labs

Tessera Labs

San Jose, CA

Research Scientist
$200k+/yrOn-siteAI Research

Research Scientist defining and executing research on reliable long-horizon agents in enterprise environments. The role focuses on post-training and reinforcement learning, agent memory, evaluation, verification, and structured representations, combining hands-on experimentation with product delivery and publication.

OpenAI

OpenAI

San Francisco, CA

People Research Scientist
$198k+/yrOn-siteAI Research

Conduct rigorous people research and applied data science to evaluate talent programs, organizational health, and employee experiences. The role requires advanced expertise in research design, experimentation, measurement, causal inference, statistical modeling, and responsible handling of sensitive employee data.

The Voleon Group

The Voleon Group

New York, NY
Member of Research Staff, Causal Inference
$250k+/yrHybridAI Research

Conducts causal inference research for financial market prediction and portfolio optimization, developing and validating models from research through live trading. Requires Ph.D.-level coursework, strong causal inference and statistics expertise, mathematical ability, and production Python skills.