Skip to content
NuroNuro

Software Reliability Engineer

Builds and operates resilient systems for autonomous vehicle fleet reliability, including pipelines for signal analysis, automated triage tools, internal workflows, and leading investigations. Requires production software experience and strong debugging skills in Python, Go, Bash, C++.

About the job

Responsibilities

  • Build fleet-scale pipelines that turn noisy onboard signals into actionable, high-confidence investigations.
  • Develop automated triage and correlation systems that deduplicate issues, route them to the right owning teams, and attach up-to-date priority signals and diagnostic context.
  • Partner with engineering teams and subject matter experts to turn investigation outcomes into better instrumentation, automation, and signal quality over time.
  • Build internal tools and workflows that reduce duplicate effort and increase situational awareness as the fleet scales (self-service debugging, standardized metrics, shared templates, securely scoped access).
  • Lead reliability investigations to identify contributing factors and ensure learnings turn into durable engineering changes.

Requirements

  • Experience writing and shipping software that runs in production, with an ownership mindset and attention to how it behaves in real-world conditions.
  • Ability to build and maintain tools and automation that enable other engineers: internal tools, instrumentation, and visualizations (Python, Go, Bash, C++).
  • Strong debugging fundamentals across the stack, including using system signals and live troubleshooting to form hypotheses and identify contributing factors.
  • Strong interest in reliability engineering as a growth path: motivated by making complex systems understandable, resilient, and easier to run as they scale.

Nice-to-Haves

  • Background in distributed systems or real-world deployed systems (vehicles, robotics, IoT, or similar).
  • Familiarity with production telemetry and observability.
  • Experience applying reliability metrics and operational feedback loops to drive improvements.
  • Exposure to cross-team reliability work in mission-critical environments.

Compensation

Base pay range: $145,830 - $219,000 (depending on experience, qualifications, education, location, skills). Eligible for annual performance bonus, equity, and competitive benefits package.

Skills

Python, Go, Bash, C++, Distributed Systems, Observability, Telemetry, Reliability Engineering, Debugging, Automation

SimplePractice

SimplePractice

United States

DevOps Engineer, Data & AI Platform
$144k+/yrOn-site3+ YOEDevOps / SRE

The DevOps Engineer will build and operate reliable infrastructure, deployment workflows, and observability for data pipelines and AI/ML systems. The role requires at least three years of DevOps, SRE, or infrastructure experience plus strong cloud, Terraform, containerization, and MLOps expertise.

Teleport

Teleport

United States

IT Security and Automation Engineer
$149k+/yrRemoteDevOps / SRE

Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.

Cloudflare

Cloudflare

Austin, TX
Systems Engineer - Database Platform
$150k+/yrHybridDevOps / SRE

Build and operate a highly available, multi-region PostgreSQL platform, developing automation, monitoring, disaster recovery, and performance tooling. Requires experience with large-scale PostgreSQL clusters, infrastructure as code, scripting, containers, and observability.

Fluidstack

Fluidstack

New York, NY
Infrastructure Deployment Engineer
$150k+/yrOn-site5+ YOEDevOps / SRE

Leads on-site deployment of data center physical infrastructure, managing contractors, performing QA/QC on fiber optics and cabling, and ensuring compliance with standards. Requires 5+ years experience, SME-level fiber optic expertise, bachelor's degree, and 40% travel readiness.

Trexquant

Trexquant

New York, NY

Python Engineer - Trade Operations
$150k+/yrOn-site3+ YOEDevOps / SRE

The Python Engineer will improve and operate trading systems, support integrations with asset classes and prime brokers, and handle monitoring, incidents, and performance optimization. The role requires 3+ years of experience, strong Python and Linux skills, and familiarity with market data and order-entry systems.