Senior Software Engineer, Infrastructure
Designs, builds, and maintains scalable infrastructure for real-time telemetry platform supporting mission-critical systems. Requires 8+ years in distributed systems, cloud environments (AWS/GCP/Azure), Kubernetes, Docker, and DevOps tools.
About the job
Responsibilities
- Design, build, and maintain scalable, resilient infrastructure solutions to support our growing platform and customer base.
- Collaborate with software engineers to optimize application performance and reliability.
- Implement monitoring, alerting, and logging systems to ensure proactive identification and resolution of issues.
- Automate deployment processes and streamline infrastructure management using modern DevOps tools and methodologies.
- Evolve our backend architecture/infrastructure for both cloud and on-premise deployments.
- Work with the team to set and prioritize our roadmap to maximize customer impact.
- Lead initiatives to improve infrastructure reliability, performance, and cost efficiency.
Requirements
- 8+ years of relevant distributed systems experience focusing on designing and managing cloud-based environments (e.g., AWS, Azure, GCP).
- Hands-on experience with containerization technologies (Docker, Kubernetes) and container orchestration platforms.
- Passion for building and operating developer productivity tools, frameworks, and other aspects of platform engineering.
- Familiarity with CI/CD pipelines and version control systems.
- Knowledge of automated testing best practices and frameworks, ensuring software reliability through integration, performance, and end-to-end testing in distributed systems.
- Experience with our tech stack: Go, Java, React, TypeScript, Kafka/Redpanda, Flink, AWS, Azure, GCP, Docker, Kubernetes, Terraform/CDK or equivalent distributed systems.
- Good knowledge of cloud, on-prem, networking, and service architecture in multi-region multi-cloud setups.
Engineering At SIFT (Relevant Technologies)
- Web frontend & backend: ECharts, Go, gRPC, PostgreSQL, Protobuf, Radix, React, Redux, TypeScript.
- Data: Arrow, DataFusion, Flink, Parquet, Rust.
- Infrastructure: Argo CD, AWS, Docker, GitHub Actions, Grafana, Kubernetes, Kustomize, Linux, Prometheus, Terragrunt.
Compensation
- Salary range: $170,000 - $220,000 per year. Plus equity and benefits.
Skills
Kubernetes, Docker, AWS, GCP, Azure, Terraform, Go, CI/CD, Prometheus, Grafana, Argo Cd, Flink, Kafka, Postgres, gRPC
Similar jobs
DevOps / SRE jobsOwn foundational cloud infrastructure and the internal developer platform supporting Commure’s engineering teams. The role requires 6+ years of infrastructure, platform, or SRE experience and hands-on expertise across Kubernetes, infrastructure as code, GitOps, observability, and cloud environments.
Leads cloud infrastructure, platform strategy, deployment pipelines, and infrastructure automation for a growing consumer platform. Requires 5+ years in infrastructure, DevOps, platform engineering, or SRE, plus deep AWS, coding, containerization, and infrastructure-as-code experience.
Own reliability, deployments, observability, compliance, and AI infrastructure across AWS and Kubernetes for a fintech platform. The role requires strong DevOps/SRE depth, backend software engineering experience, and hands-on ownership of SOC 2 and PCI-DSS controls.
Own and evolve secure, highly available AWS and Azure infrastructure, including Terraform automation, Kubernetes, CI/CD, observability, networking, and incident response. The role requires 7+ years of DevOps or related experience and strong cross-functional partnership across engineering and security.
Own and evolve VSCO’s AWS/EKS platform, including infrastructure as code, GitOps, CI/CD, observability, networking, and production reliability. The role requires 5+ years of hands-on infrastructure or SRE experience and strong Kubernetes, Terraform, and AWS expertise.