Research Internship
Conducts and publishes cutting-edge machine learning research, building and training large language models and contributing to applied AI product initiatives. Applicants should be pursuing a PhD or demonstrate exceptional equivalent experience, with expertise in ML systems, Transformers, programming, and modern ML frameworks.
About the job
Responsibilities
- Conduct cutting-edge machine learning research, building and training large language models.
- Focus on research projects expanding the frontier of language modeling, including evaluation, multimodal models, and optimization.
- Disseminate research results through publications, datasets, and code.
- Contribute to research initiatives with practical applications in product development.
Requirements
- Currently pursuing, or in the process of obtaining, a PhD in Machine Learning, NLP, Artificial Intelligence, or a related discipline; exceptional non-PhD candidates may also be considered.
- Available for a full-time internship lasting 4–6 months.
- Eligible for work authorization in the country of employment throughout the internship.
- Experience with large-scale distributed training strategies, data annotation and evaluation pipelines, or state-of-the-art ML models.
- Familiarity with autoregressive sequence models such as Transformers.
- Strong communication and problem-solving skills.
- Knowledge of Python, C, C++, Lua, or related programming languages.
- Knowledge of JAX, PyTorch, and TensorFlow.
- Experience building systems using machine learning and deep learning techniques.
- Passion for applied NLP models and products.
Preferred Qualifications
- Publications in top-tier venues across machine learning, NLP, artificial intelligence, computer vision, optimization, computer science, statistics, applied mathematics, or data science.
- Ability to tackle analytical problems using quantitative methodologies.
- Proficiency handling and analyzing complex, high-dimensional data from various sources.
- Experience applying theoretical and empirical research to real-world problem-solving.
Benefits
- Weekly lunch stipend of $75/£75 or equivalent in local currency.
- Full health and dental benefits, including a separate mental health budget.
- RRSP matching, 401(k), or pension scheme.
- Parental leave top-up for up to 6 months for either parent.
- Annual enrichment benefits for arts and culture, fitness and wellness, quality time, and workspace improvements.
- Education and learning stipend for conferences, courses, and coaching.
- Six weeks of paid vacation.
- Travel budget for remote employees visiting other offices and an annual company offsite.
- Co-working benefit for employees not near an office.
- $500 home office stipend.
Skills
Machine Learning, Natural Language Processing, Artificial Intelligence, LLMs, Transformers, Python, C++, JAX, PyTorch, TensorFlow, Distributed Training, Deep Learning, Computer Vision, Data Evaluation
Similar jobs
AI Research jobsSummer 2027 internship applying computer science, mathematics, statistics, and machine learning research to practical product capabilities. The intern will prototype, evaluate, and communicate solutions involving data privacy, security, governance, algorithms, and text analytics.
Conduct foundational research on LLMs and multimodal systems, designing architectures and training methods and helping move prototypes into production. The role targets PhD researchers graduating by December 2026 with strong machine-learning research and programming experience.
AI Research Intern researching agentic AI applications for customer-facing products and developing working prototypes. The role requires current pursuit of a technical bachelor's degree, prior software engineering or substantial project experience, and interest in LLMs or generative AI.
Paid internship for quantitative students contributing to AI-powered software creation, systems optimization, and developer tooling. Candidates should demonstrate strong mathematical ability, curiosity about AI, and autonomous cross-functional collaboration.
Researcher developing and publishing mechanistic interpretability techniques, building infrastructure to study model internals, and guiding alignment-focused research. Requires research experience in machine learning or a related field, strong engineering skills, and proficiency in Python or similar languages.