Latest ML Engineering jobs at Pulse
Job results
Develops low-latency, high-throughput inference services for OCR and multimodal models, optimizing batching, kernels, and autoscaling while evaluating serving frameworks. Requires 3+ years in performance engineering or ML systems with strong Python and GPU experience.
Develops and fine-tunes specialized vision and language models for document understanding, including OCR, layout, tables, and charts. Requires 3+ years ML experience with PyTorch/JAX and strong engineering focus; onsite in San Francisco.