Applied AI/ML Engineer
Builds state-of-the-art document processing infrastructure using LLMs, including QA agents, optimizers, multimodal models, and self-correcting systems. Monitors production models, runs experiments, and owns large product areas for real-world customer impact.
About the job
Responsibilities
- Design novel LLM techniques for increasing the complexity of use cases and data streams that Extend can be applied to.
- Monitor existing models in production, understand what’s working well (and what isn’t), and run experiments to solve those issues.
- Build a QA Agent to flag low confidence results.
- Deploy an optimizer agent in a loop that improves document performance in the background.
- Create multimodal models for document layout awareness.
- Develop novel chunking strategies for handling long documents.
- Build data pipelines and evaluate model performance.
- Build a self-correcting system that automatically gets better over time.
- Have complete ownership over the work you do and direct relationship with customers.
Benefits
- 90% health insurance premium coverage.
- Relocation for candidates based outside of NYC.
- Unlimited PTO policy.
- Daily lunch & snacks covered (plus dinner if late).
- Unlimited token / tooling access.
- Learning and development investment.
Skills
LLMs, Machine Learning, AI, Python, PyTorch, Transformers, Multimodal Models, Data Pipelines, Production Ml, Llm Agents
Similar jobs
ML Engineering jobsBuild and operate large-scale ranking and retrieval systems that power search relevance, including hybrid lexical/vector search, embeddings, query understanding, and permission-aware retrieval. Requires a bachelor's degree and 5+ years of ML engineering experience in ranking or information retrieval.
Build and operate production ML infrastructure spanning training, deployment, serving, monitoring, data pipelines, and feedback-driven retraining. The role requires strong MLOps and DevOps experience, Python and SQL proficiency, and ownership of reliable cloud-based systems.
Build and scale post-training, reinforcement-learning, evaluation, and inference systems for long-horizon agents operating over complex enterprise software. The role requires strong Python and PyTorch or JAX skills, distributed GPU experience, empirical rigor, and the ability to take research results into production.
Build and operate production AI agents that transform enterprise processes, data, and code. The role focuses on tool layers, retrieval, context management, evaluations, monitoring, auditability, and guardrails, requiring strong Python and TypeScript plus experience with production LLM systems and traditional machine learning.
Build and productionize applied AI/ML systems for document understanding, agentic workflows, and demand forecasting using rich, messy enterprise data. The role requires 3+ years of production AI/ML experience, strong evaluation and monitoring practices, and a STEM master’s degree.