Software Engineer, AI Inference / HPC
Develops and optimizes AI inference engine for image/video enhancement, focusing on performance, GPU/CPU optimization, model deployment, and hardware partnerships. Requires C/C++ expertise, 1+ years experience in performance optimization and image processing.
About the job
Responsibilities
- Improve performance, stability, availability of new features, and simplify/improve the API of the internal AI Engine framework.
- Act as technical bridge between Deep Learning research team and Production products.
- Prepare new & updated models for production.
- Optimize GPU/CPU for inference.
- Work with hardware partners (NVIDIA, AMD, Intel, Apple) to optimize inference on their hardware.
Requirements
- Hands-on experience with performance optimization (concurrency, multithreading, memory, speed, benchmarking, reliability).
- Experience architecting APIs for internal development.
- Hands-on experience implementing image processing or computational photography algorithms.
- Expert knowledge of C/C++.
- At least 1+ years of professional working experience in a related field.
Preferred
- Experience with video encoding/decoding and file formats.
- Experience with OpenCV, ffmpeg, GPU programming.
- Experience with raw image camera pipeline and image formats.
- Experience with ONNX, CoreML, TensorRT runtime SDKs.
- Interest in photography or videography.
Skills
C++, C, Performance Optimization, Multithreading, Concurrency, API Design, Image Processing, Opencv, Ffmpeg, Gpu Programming, Onnx, Coreml, TensorRT
Similar jobs
ML Engineering jobsBuild and maintain tooling, evaluation systems, quality gates, and infrastructure for MongoDB's agent skills and AI platform. Requires 2+ years building production software, developer tools, CLIs, test infrastructure, or CI/CD, with strong fundamentals in API design, testing, and reasoning about nondeterministic AI systems.
Develop and deploy machine learning models for biomedical research and AI product development, collaborating with scientific, engineering, and product teams. Requires a master's degree with 2–4 years of experience or a PhD with 0–2 years, plus strong Python and ML development skills.
Build and deploy production voice AI agents for customers, creating demos, debugging edge cases, improving performance, and translating feedback into product improvements. The role combines hands-on engineering, customer engagement, and pre- and post-sales delivery.
Research and develop computer vision and deep learning algorithms for autonomous drones, taking ownership of projects from prototyping through product integration. The internship requires strong C++ or Python and PyTorch skills, mathematical foundations, and software engineering ability.
Builds high-scale data pipelines, distributed systems, and AI agent workflows using LLMs for fraud intelligence platform. Requires 2+ years software engineering, Python proficiency, big data tools, AWS/K8s, and ML foundations.