Research Engineer, Retrieval & Search, Applied Engineering
Develop and deploy retrieval and search algorithms for OpenAI's API and ChatGPT, collaborating with research teams on production systems for millions of users. Requires experience with ML systems, vector databases, and large-scale search.
About the job
Responsibilities
- Work on retrieval & search algorithms and methodologies in close collaboration with our research team, including problems in such domains as document search, enterprise search, knowledge retrieval, and web-scale search.
- Deploy these search methodologies into production in both the API and ChatGPT to be used by millions of end users.
- Explore novel research topics in retrieval & search that may inform our product strategy in the medium and long term.
- Partner with researchers, engineers, product managers, and designers to bring new features and research capabilities to the world.
Requirements
- Extensive prior experience building and maintaining production machine learning systems.
- Prior experience working with vector databases, search indices, or other data stores for search and retrieval use cases.
- Prior experience building and iterating on internet-scale search systems.
- Own problems end-to-end, and are willing to pick up whatever knowledge you're missing to get the job done.
- Have the ability to move fast in an environment where things are sometimes loosely defined and may have competing priorities or deadlines.
Skills
Machine Learning, Vector Databases, Search Indices, Retrieval Algorithms, Production Ml Systems, Internet-Scale Search, Document Search, Enterprise Search, Knowledge Retrieval, Web-Scale Search
Similar jobs
ML Engineering jobsBuild and optimize OpenAI’s inference stack for AWS Trainium across high-performance kernels, compilers, runtimes, and model execution. The role requires systems programming and accelerator experience, with opportunities to solve end-to-end performance problems for frontier-scale AI models.
Build and deploy LLM-powered tools, agents, and ecosystem infrastructure with life sciences research institutions. The role requires deep scientific or biomedical research experience, production software development expertise, and the ability to translate partner workflows into scalable AI systems.
Build research infrastructure and tooling that enables AI models to design silicon, including reinforcement learning environments, EDA integrations, evaluations, and experiment workflows. The role requires strong software engineering fundamentals and comfort working across research, tooling, and chip-design systems.
Build and optimize the production LLM inference runtime for frontier models on OpenAI’s custom silicon. The role spans scheduling, distributed execution, memory and KV-cache management, performance tooling, and hardware-software co-design.
Build and operate machine learning models for sales roleplay, scoring, and coaching products, owning the lifecycle from fine-tuning and evaluation through production and on-device deployment. The role emphasizes open-source models, latency and privacy optimization, and rigorous model testing.