Build, optimize, and maintain large-scale build systems (Bazel priority) and developer infrastructure including CI/CD, observability, and release automation for a fast-growing hardware-software platform company. Requires 4+ years experience with build systems at scale and large monorepos.
130k – 230k/yr
On-site4+ YOEDevOps / SRE
About the role
Build, optimize, and maintain Nominal's large-scale build systems, with Bazel as the highest priority.
Improve CI/CD pipeline performance, release automation, and release reliability.
Own developer infrastructure that enables engineering velocity as our monorepo and engineering organization scale.
Build and maintain observability, telemetry, logging, and metrics infrastructure using tools like Datadog, Grafana, and OpenTelemetry.
Partner closely with product engineering teams to remove developer friction and improve day-to-day engineering productivity.
Skills
4+ years building, optimizing, and maintaining build systems at scale.
Experience with Bazel or similar build systems.
Experience working in and scaling large monorepos.
A track record of improving CI/CD performance, release reliability, and developer productivity using measurable outcomes.
Practical, startup-minded engineering philosophy - You don't let perfect be the enemy of good, but you approach problems from first principles and build solutions that scale.
Comfortable owning developer infrastructure end to end, from build tooling and CI/CD to release automation and engineering productivity.
Nice to have
Datadog
Grafana
OpenTelemetry
Audit logging
Kubernetes
Remote build caches
Ephemeral development environments, or other developer platform tooling.
Benefits
100% coverage of medical, dental, and vision insurance
Unlimited PTO and sick leave
Free lunch, snacks, and coffee
Professional Development Stipend
In-office hardware lab with a $250 project stipend
Annual company retreat
Compensation
The base pay range for this role is $130,000 – $230,000 per year.
Build and maintain an internal agentic AI platform to accelerate developer workflows at Tulip. Identify high-impact AI opportunities in code generation, testing, debugging and tooling; own evals, standards, onboarding and measurement of AI adoption. Requires 5+ years software engineering experience with strong hands-on LLM/agentic AI and full-stack TypeScript skills.
130k – 180k/yrHybrid5+ YOEDevOps / SRE
Site Reliability Engineer III
OnxmapsBozeman, MT
Site Reliability Engineer responsible for deploying, monitoring, and maintaining highly available infrastructure on GCP using Terraform, Kubernetes, and various cloud services. Requires 5+ years experience (3+ in production), strong Kubernetes/IaC background, and on-call participation to ensure reliable systems for millions of users.
130k – 153k/yrHybrid5+ YOEDevOps / SRE
Software Engineer, Infrastructure
KustomerNew York, NY
Infrastructure Software Engineer building scalable backend systems, observability, and developer tools on the Foundation team. Lead projects on database sharding, event bus, search scaling, and latency; mentor engineers. Requires 5+ years with distributed systems, NoSQL (MongoDB), IaC (Terraform), and architecture ownership.
130k – 215k/yrHybrid5+ YOEDevOps / SRE
Linux Systems Engineer (USA)
TrexquantStamford, CT +1
Hands-on Linux Systems Engineer builds and maintains bare-metal servers, manages storage like ZFS, automates with Ansible and Bash, and ensures production reliability. Requires 3+ years Linux experience, physical server management, and on-call rotation with data center travel.
Designs, deploys, and manages HashiCorp Vault clusters for secure secret management in on-premises and cloud (AWS/GCP) hybrid environments with Kubernetes integration. Requires 3+ years experience, zero trust principles, IaC tools like Terraform, and automation scripting.