Software Engineer, Kernel Performance & AI Tooling
Develops kernel performance optimizations, AI-assisted tooling, and observability infrastructure for AI-native hardware. Requires strong low-level systems experience, kernel/accelerator expertise, and familiarity with AI workflows for engineering acceleration.
About the job
Responsibilities
- Build developer tooling and workflows that make kernel development and performance optimization faster, more scalable, and easier to debug, integrate, and deploy.
- Develop observability, diagnostics, and validation infrastructure that makes AI-assisted optimization systems more interpretable, reliable, and effective.
- Optimize production kernels end to end by formulating optimization problems, running search loops, analyzing bottlenecks, debugging generated implementations, and landing improvements into production.
- Design abstractions, interfaces, and automation systems that accelerate kernel optimization, correctness validation, and hardware-software co-design.
- Improve AI-assisted optimization systems for specialized tasks through better datasets, evaluations, benchmarking, and research infrastructure.
- Partner across research and engineering teams to turn new ideas into practical systems spanning production needs and long-term infrastructure strategy.
Requirements
- Strong systems or tooling engineering experience, with a background in low-level software, performance optimization, or infrastructure.
- Experience with developer tooling, debugging infrastructure, profiling, observability, or workflow design for technical users.
- Depth in kernel development, accelerator architecture, compiler systems, or related performance-critical domains.
- Familiarity with AI-assisted systems, agentic workflows, post-training, or reinforcement learning for engineering or research applications.
- Strong experimental judgment, comfort with ambiguity, and the ability to move fluidly between research exploration and production execution.
- Interest in compilers, DSLs, program synthesis, or AI for systems.
Preferred
- Strong systems and tooling engineer with real depth in kernels and accelerators.
- Comfortable working across software and hardware boundaries, can reason deeply about performance, abstractions, and system design.
- Hands-on experience optimizing code for GPUs, high-performance CPUs, or custom accelerators.
- View AI not as the end product, but as a force multiplier for engineering productivity and system optimization.
Skills
Kernel Development, Performance Optimization, Gpus, Cpus, Compilers, Developer Tooling, Observability, Profiling, Ai-Assisted Systems, Reinforcement Learning, Dsls, Program Synthesis, Accelerator Architecture, Hardware-Software Co-Design, Debugging Infrastructure
Similar jobs
DevOps / SRE jobsBuild and operate the Kubernetes-based cloud and on-premises infrastructure powering large-scale crawling, search, and ML workloads. The role requires 5+ years in DevOps, platform engineering, or cloud infrastructure, with strong Kubernetes, cloud, Docker, Terraform, and distributed-systems experience.
Owns reliability standards, incident management, observability, failure testing, and automation for a high-throughput AI infrastructure platform. The role requires deep Linux, networking, software, cloud-native, and distributed-systems experience, along with the ability to influence teams across the organization.
Build and operate scalable build systems, CI pipelines, and developer infrastructure for consumer-device software. The role requires 5+ years of engineering experience, expertise with Bazel or comparable build systems, and experience improving CI reliability and performance at scale.
Designs, operates, and improves secure enterprise networks spanning offices, campuses, cloud environments, and connectivity services. The role combines architecture, production operations, troubleshooting, observability, security, and infrastructure automation.
Build and operate distributed infrastructure across compute, storage, networking, data, deployment, and reliability domains. The role requires 4+ years of backend or platform engineering experience, strong systems-language skills, and the ability to own complex production systems and lead cross-team technical initiatives.