Senior Software Engineer, Cloud Infrastructure
Senior engineer owning reliability, automation, and evolution of core cloud infrastructure systems. Builds tooling, integrations, and observability; troubleshoots across backend, infra, and frontend. Requires Rust proficiency, database expertise, and distributed systems experience.
About the job
What You’ll Do
- Design, implement, and maintain core systems and services that support our platform’s reliability, performance, and scalability.
- Build automation and tooling that reduce operational overhead, improve developer workflows, and enforce system-level consistency.
- Develop integrations between our systems and external platforms (e.g., billing, CRM, authentication, analytics).
- Improve observability, monitoring, and alerting across services to ensure strong operational visibility.
- Architect and maintain infrastructure components (e.g., distributed services, data pipelines, deployment automation).
- Contribute to security-focused systems work, such as permissions controls, access flows, and auditability.
- Troubleshoot issues across services and layers, backend, infrastructure, and occasionally frontend, taking full ownership from diagnosis to resolution.
- Jump into frontend code as needed to close the loop on system-level changes or fix issues that block broader system reliability.
What You’ll Need
- B.S. in Computer Science or a related field, or equivalent experience.
- Strong professional experience in systems, platform, or infrastructure engineering.
- Proficiency in Rust with a solid understanding of systems-level programming.
- Solid experience with databases (transactional & analytical) including schema design, performance tuning, migrations, and operational reliability.
- Strong understanding of API design, data flows, and service integrations.
- Hands-on experience with distributed systems or service-oriented architectures.
- Comfort working with infrastructure-as-code and containerized deployments.
- Ability to work autonomously and make high-leverage technical decisions in a fast-moving environment.
- A bias toward automation and building tools that remove friction.
Nice to Have
- Experience with frontend development (JavaScript/React) for occasional UI-level fixes.
- Background with billing systems (Stripe) or CRM platforms.
- Experience building internal tooling, permissions systems, or developer productivity infrastructure.
- Familiarity with CAD-adjacent workflows or domain-specific data pipelines.
Compensation
Salary Range: $145,000—$195,000 USD. In addition to salary, competitive equity and full benefits package.
Skills
Rust, Databases, API Design, Distributed Systems, Infrastructure As Code, Containerization, Observability, Monitoring, Alerting, Stripe
Similar jobs
DevOps / SRE jobsSenior Site Reliability Engineer responsible for operating and improving large-scale, FedRAMP-compliant cloud services through automation, observability, incident response, and platform engineering. Requires strong Kubernetes, cloud infrastructure, software engineering, and reliability engineering expertise.
The Senior Site Reliability Engineer will build and operate secure, highly available infrastructure and Snowflake data tooling for large-scale SaaS systems. The role emphasizes automation, Kubernetes, Terraform, CI/CD, incident response, and collaboration with development, data science, and security teams.
The Senior Site Reliability Engineer will build and operate secure, scalable infrastructure and Snowflake data systems, automate deployments and operational processes, and lead incident response. The role requires strong coding, Terraform, Kubernetes, CI/CD, and data-platform experience, plus U.S. Person status.
Designs and operates shared cloud and private-cloud platforms, infrastructure automation, Kubernetes capabilities, and developer self-service tools. Requires 7+ years in platform, cloud infrastructure, DevOps, or SRE, with strong Terraform, Ansible, Linux, Kubernetes, and public-cloud experience.
Designs, deploys, and operates secure, resilient enterprise and cloud networks across data centers, on-premises environments, and AWS and Azure. Requires 6+ years of production network experience plus expertise in routing, switching, firewalls, automation, and hybrid connectivity.