AI Engineer
Build and deploy production GenAI/LLM applications and intelligent agents to automate workflows across Procurement, Supply Chain, Legal, Finance, HR, and Marketing. Requires 5+ years engineering experience including 2+ years production LLMs, Python, LangChain/LlamaIndex, and cloud AI/vector DB tools.
About the job
Responsibilities
- Design and implement production-grade LLM applications, managing the full stack from data ingestion and vector database integration to prompt engineering, fine-tuning, and model evaluation.
- Build and deploy sophisticated intelligent agents capable of complex reasoning, secure tool usage, and autonomous execution of multi-step business workflows.
- Own the deployment process, including self-healing systems, latency optimization, cost management, and robust performance monitoring and alerting in a large-scale enterprise environment.
- Partner closely with cross-functional teams (Legal, Finance, HR, etc.) to identify operational bottlenecks and translate them into efficient, code-driven AI automation.
- Implement rigorous standards for security, data privacy, and accuracy, ensuring the AI framework integrates seamlessly with existing corporate infrastructure.
Requirements
- 5+ years in Data Engineering, Software Engineering, or Data Science, with at least 2+ years of hands-on experience deploying GenAI/LLMs in a production enterprise environment.
- Deep proficiency in Python and frameworks such as LangChain, LlamaIndex, or AutoGen.
- Experience with cloud AI services (AWS Bedrock or Google Vertex AI) and vector databases (Pinecone, Weaviate, Milvus, or OpenSearch).
- Experience building agents and AI workflows.
- Proven ability to translate business pain into technical requirements.
Nice-to-Haves
- Previous experience building AI solutions for corporate functions like Finance, Legal (e.g., contract analysis), HR (e.g., policy retrieval), or Supply Chain (e.g., forecasting efficiency).
- Familiarity with automated evaluation pipelines like RAGAS, Arize, or other LLM-based evaluation metrics.
- Experience with techniques such as quantization, speculative decoding, or efficient fine-tuning (LoRA/QLoRA) to improve model performance and reduce inference costs.
- Previous experience in the autonomous vehicle, robotics, or high-tech manufacturing sectors.
Skills
Python, LangChain, Llamaindex, Autogen, Aws Bedrock, Google Vertex Ai, Pinecone, Weaviate, Milvus, Opensearch, LLMs, RAG, Generative AI
Similar jobs
ML Engineering jobsBuild production-grade AI agents, evaluation infrastructure, and developer tooling that make AI-assisted engineering faster, safer, and reusable across teams. The role requires software engineering experience, platform or internal developer-product experience, and hands-on expertise with LLM integration and orchestration.
Build trustworthy infrastructure for production LLM agents, closed-loop evaluation, and autonomous research workflows. The role requires strong Python and distributed-systems experience, hands-on LLM post-training and inference knowledge, and experience operating agent systems at scale.
Build and operate large-scale ranking and retrieval systems that power search relevance, including hybrid lexical/vector search, embeddings, query understanding, and permission-aware retrieval. Requires a bachelor's degree and 5+ years of ML engineering experience in ranking or information retrieval.
Build and operate production ML infrastructure spanning training, deployment, serving, monitoring, data pipelines, and feedback-driven retraining. The role requires strong MLOps and DevOps experience, Python and SQL proficiency, and ownership of reliable cloud-based systems.
Build and scale post-training, reinforcement-learning, evaluation, and inference systems for long-horizon agents operating over complex enterprise software. The role requires strong Python and PyTorch or JAX skills, distributed GPU experience, empirical rigor, and the ability to take research results into production.