Senior Software Engineer - Platform & Infrastructure
Founding Senior Platform Engineer building and owning AWS cloud infrastructure, reliability, observability, security/compliance (SOC 2, Vanta), and release tooling for a fintech platform serving banks and credit unions.
About the job
Responsibilities
- Own and evolve the reliability posture of the ModernFi network (observability, incident response, SLOs, and on-call practices)
- Lead the buildout of a more advanced release pipeline with merge queues, feature-flag tooling, and per-domain release gates for highest-stakes services
- Take ownership of cloud platform, Pulumi modules, AWS account structure, identity, networking, and tooling for product engineers
- Drive security and compliance posture: Vanta remediation, SOC 1/2 readiness, vulnerability scanning, secrets and key management, and hardening across the stack
- Upgrade on-call tooling, runbooks, rotations, and postmortems while participating in infrastructure on-call rotation
- Partner with product engineering pods to ship abstractions that reduce platform drag
- Cultivate operational excellence through code reviews, design discussions, and mentorship
Requirements
- 5+ years of software engineering experience with significant time on infrastructure, platform, or reliability-focused work in production
- Deep fluency with cloud infrastructure (AWS preferred), infrastructure-as-code (Pulumi, Terraform), and modern CI/CD pipelines
- Security mindset with familiarity in SOC 2, compliance tooling like Vanta, secrets management, key rotation, and least-privilege access
- Hands-on experience with modern observability stack (Datadog, Prometheus + Grafana, Elastic stack) and track record of setting observability standards, defining SLOs, running postmortems
- Comfortable owning systems end-to-end from IaC provisioning to recovery runbooks
- Experience in small, fast-moving engineering teams with track record of improving reliability
- Excellent communication skills and ability to explain complex concepts clearly
Tech Stack
- AWS / Pulumi IaC
- PostgreSQL (RDS), ECS, Docker
- Temporal
- Python (FastAPI, SQLAlchemy)
- Datadog, PagerDuty, GitHub Actions
- Auth0, Vanta, Fern
Skills
AWS, Pulumi, Terraform, Python, FastAPI, Sqlalchemy, Docker, ECS, Datadog, Postgres, GitHub Actions, Temporal, Observability, CI/CD, SOC 2
Similar jobs
DevOps / SRE jobsSenior software engineer responsible for operating and evolving Voltus’s infrastructure platform across AWS, Kubernetes, Nomad, observability, stateful systems, and developer tooling. The role requires 6+ years of engineering experience, deep production Kubernetes and AWS expertise, and strong Go or Python skills.
Own and evolve VSCO’s AWS/EKS platform, including infrastructure as code, GitOps, CI/CD, observability, networking, and production reliability. The role requires 5+ years of hands-on infrastructure or SRE experience and strong Kubernetes, Terraform, and AWS expertise.
Build and operate highly available, distributed platform services and cloud infrastructure for petabyte-scale observability products. The role requires 6+ years of experience, strong Java and AWS expertise, Kubernetes and Terraform production experience, and a bachelor’s degree or equivalent.
Own foundational cloud infrastructure and the internal developer platform supporting Commure’s engineering teams. The role requires 6+ years of infrastructure, platform, or SRE experience and hands-on expertise across Kubernetes, infrastructure as code, GitOps, observability, and cloud environments.
The Senior Network Engineer will design, automate, and operate large-scale, high-performance network infrastructure for AI data centers and GPU clusters. The role requires 5+ years of data center networking experience, expertise in spine-leaf fabrics and routing protocols, and familiarity with HPC or GPU-dense environments.