Skip to content

Platform Engineer

Build foundational infrastructure platforms and operational tooling for global data center operations, including CMDB, DCIM, asset management, automation, and observability systems. Requires 3+ years of production software development experience and strong Python or Go skills.

About the job

Responsibilities

Infrastructure Platform Development

  • Design and build a next-generation CMDB as the authoritative source for infrastructure assets, network topology, and configuration data.
  • Create DCIM platforms for rack operations, server/GPU deployment, OS installation, quality assurance, and white-screen operations.
  • Build asset lifecycle management systems for receiving, racking, inventory, break-fix, and decommissioning workflows.
  • Develop monitoring and observability platforms integrating telemetry from BMS, EPMS, and IT devices with intelligent alarming and incident management.
  • Create self-service portals and automation for regional bootstrap, day-two operations, and fleet-scale management.

Operational Excellence and Automation

  • Automate manual workflows and build self-service tools for operations and engineering teams.
  • Develop workflow orchestration systems spanning incident, problem, and change management.
  • Build digital-twin visualizations and operational dashboards; partner with data teams on analytics.
  • Create integration layers connecting internal platforms with external vendors and third-party systems.

Technical Leadership and Reliability

  • Collaborate with data center operations, systems engineering, network engineering, security, product, support, and business stakeholders.
  • Evaluate build-versus-buy decisions for platform components.
  • Champion CI/CD, infrastructure as code, automated testing, and observability-first development.
  • Participate in architecture reviews, code reviews, documentation, and knowledge sharing.
  • Design high-performance, fault-tolerant systems capable of handling thousands of QPS.
  • Implement monitoring, logging, debugging, error handling, data migration strategies, and dependency management.
  • Own projects end to end through deployment and production readiness.

Requirements

  • 3+ years of professional software development experience building production systems.
  • Strong programming skills in Python, Go, or similar languages, with knowledge of system design patterns.
  • Experience designing RESTful APIs, data models, and distributed systems.
  • Proficiency with relational and NoSQL databases.
  • Hands-on experience with Docker and infrastructure-as-code tools.
  • Understanding of CI/CD pipelines and modern development workflows.
  • Knowledge of TCP/IP, DNS, HTTP, and Linux/Unix environments.
  • Strong problem-solving, communication, scalability, reliability, and operational skills.
  • Bachelor's degree in Computer Science or equivalent practical experience.

Nice-to-Haves

  • Experience with CMDB systems such as NetBox or Device42, or asset management platforms.
  • Background in infrastructure automation, DevOps, or platform engineering.
  • Familiarity with Temporal, Airflow, or Camunda.
  • Knowledge of Prometheus, Grafana, or OpenTelemetry.
  • Experience with time-series databases and data visualization.
  • Understanding of ITSM frameworks such as ITIL.
  • Experience in data center operations, facilities management, or physical infrastructure.
  • Contributions to open-source infrastructure projects.

Compensation and Benefits

  • Salary: $224,000–$279,000 annually, plus potential equity in the form of restricted stock units.
  • Retirement or pension plan in line with local norms.
  • Health, dental, and vision insurance.
  • Generous paid time off policy in line with local norms.

Skills

Python, Go, REST APIs, Distributed Systems, Postgres, Redis, Docker, Terraform, Ansible, CI/CD, Linux, Prometheus, Grafana, OpenTelemetry

Rain

Rain

Remote

Software Engineer - Wallets
$220k+/yrRemoteBackend Engineering

Design and lead secure, scalable wallet infrastructure spanning custody, key management, signing, authorization, and recovery. The role requires deep expertise in wallet or cryptographic security infrastructure, distributed systems architecture, and modern blockchain account models.

Perplexity

Perplexity

San Francisco, CA

Member of Technical Staff
$220k+/yrOn-site4+ YOEBackend Engineering

Build and scale distributed backend systems powering Perplexity’s realtime voice and multimodal AI experiences. The role requires 4+ years of backend or distributed-systems experience and strong Rust, Python, or Go skills, with opportunities to work across SDKs, infrastructure, models, and agent orchestration.

OpenAI

OpenAI

San Francisco, CA

Software Engineer, Financial Engineering
$230k+/yrHybrid5+ YOEBackend Engineering

Build and operate backend and financial infrastructure supporting pricing, payments, billing, subscriptions, entitlements, and invoicing. The role requires 5+ years of software engineering experience with distributed systems, transactional workflows, and highly reliable, auditable platforms.

Anyscale

Anyscale

San Francisco, CA
Software Engineer
$215k+/yrOn-site5+ YOEBackend Engineering

Develop and improve Ray Core’s C++ distributed-systems backend, focusing on performance, reliability, fault tolerance, and scalability. The role requires at least five years of experience with distributed systems, C/C++, low-level operating systems, algorithms, and system design.

Stripe

Stripe

Seattle, WA

Software Engineer
$235k+/yrHybrid5+ YOEBackend Engineering

Build and operate scalable backend services, APIs, and real-time data systems while contributing to architecture, reliability, and product-driven solutions. Requires a relevant master’s degree with 3 years of experience or a bachelor’s degree with 5 years, plus broad distributed systems and programming expertise.