AI Engineer, Multimodal LLMs
Builds and optimizes multimodal LLM-based AI agents for enterprise conversations, integrating with systems via APIs and automating high-stakes workflows. Requires 3+ years in AI engineering, Python/PyTorch proficiency, and experience with LLMs or vision models.
About the job
Responsibilities
- Build, deploy, and optimize AI agents that engage in enterprise-grade conversations.
- Design & develop next-gen multimodal LLM architectures (LLMs, speech, vision, reinforcement learning).
- Explore optimal trade-offs between model quality and efficiency when translating research into practical solutions.
- Refine training paradigms for real-world applications.
- Integrate AI agents with enterprise systems via APIs, databases, and automation tools.
- Experiment rapidly to improve AI-driven interactions, response quality, and automation capabilities.
- Collaborate with cross-functional teams (engineering, research, and product) to shape Eloquent AI’s roadmap.
- Monitor and improve agents’ performance via user simulations and evaluations.
Requirements
- 3+ years of experience in software development, AI engineering, or NLP in a production environment.
- Strong proficiency in Python, with experience in frameworks like PyTorch and TensorFlow.
- Experience working with LLMs or large computer vision models, or generative AI models, including fine-tuning and inference optimization.
- Familiarity with APIs, cloud infrastructure (AWS, GCP, or Azure), and enterprise integrations.
- Ability to prototype, experiment, and iterate quickly to improve AI agents.
- Strong problem-solving skills and the ability to work closely with customers to refine AI solutions.
- Solid mathematical foundation of machine learning and deep learning techniques.
Nice-to-Haves
- Experience with prompt engineering, parameter-efficient fine-tuning (PEFT), retrieval-augmented generation (RAG), reinforcement learning for LLMs.
- Published AI research in top tier AI conferences like: NeurIPS, ACL, SIGIR, ICML and ICLR.
- Contributed to open-source NLP projects.
- Worked in a fast-paced startup environment and thrive in rapid iteration cycles.
Skills
Python, PyTorch, TensorFlow, LLMs, Fine-Tuning, Inference Optimization, AWS, GCP, Azure, Prompt Engineering, Peft, RAG, Reinforcement Learning, Computer Vision, NLP
Similar jobs
ML Engineering jobsBuild customer-facing integrations and evaluation workflows for leading AI labs while shipping pragmatic full-stack solutions. The role requires 4+ years of software engineering experience, TypeScript, web frameworks, LLM APIs, and cloud or CI/CD familiarity.
Build and ship AI-powered agents, automations, dashboards, and integrations that improve GPU capacity operations across Compute and C3. The role requires 3+ years of AI automation or technical operations experience, production Vercel expertise, agent-tool fluency, and strong API integration skills.
Build and optimize a high-scale LLM inference engine spanning accelerator programming, host-device coordination, and distributed systems. The role requires strong systems programming, performance analysis, and an understanding of LLM inference across compute, memory, and interconnects.
Build and operate production AI agents, automation workflows, and integrations that improve complex business processes. The role requires 5+ years of software engineering experience, modern LLM and agent-framework expertise, systems integration skills, and strong cross-functional collaboration.
Build and operate the engineering systems that support post-training research, including reinforcement learning infrastructure, sandboxed execution, data pipelines, and agent scaffolding. The role requires strong Python and systems engineering skills, project ownership, and a relevant bachelor’s degree or equivalent experience.