Software Engineer, Network Infrastructure
Design, build, and operate Render's core networking stack across data centers and clouds, focusing on Kubernetes and Linux internals, traffic routing, and hybrid connectivity at scale.
About the job
What You'll Do
- Own Render's core network infrastructure across multiple data centers and cloud providers.
- Shape how networking evolves as Render rapidly scales.
- Design and build customer-facing networking capabilities that give users greater flexibility in how their services connect, communicate, and how traffic is routed.
- Investigate complex networking issues across the stack, from the kernel and data plane to distributed systems and edge networking.
- Drive performance and reliability improvements through systematic profiling and tuning.
- Partner with engineers across the company to build a platform that is stable, predictable, and secure.
- Participate in our on-call rotation and help continuously improve how we detect, respond to, and learn from incidents.
What We're Looking For
- At least 6 years of experience building and operating large-scale networking or infrastructure systems.
- Experience with Kubernetes networking internals (e.g. kube-proxy, CNI plugins, service routing) or similar container networking technology.
- Experience with Linux networking internals (e.g. routing, netfilter/iptables/nftables, conntrack) in production environments.
- Experience building and operating traffic routing systems (e.g. load balancers, ingress, proxies) at scale.
- Strong experience designing, debugging, and operating distributed systems.
- Experience planning and executing high-risk changes with minimal downtime.
Nice-to-Haves
- Experience with data center or hybrid cloud networking (e.g. BGP, AWS Direct Connect, Arista/Juniper).
- Experience with eBPF or advanced Linux kernel networking.
- Experience developing systems in Go or similar languages that interact with low-level OS and networking primitives.
Skills
Kubernetes Networking, Kube-Proxy, Cni Plugins, Linux Networking, Iptables, Nftables, Conntrack, Load Balancing, Ingress, Proxies, Distributed Systems, BGP, Ebpf, Go
Similar jobs
DevOps / SRE jobsBuild and improve cloud infrastructure, developer workflows, and internal tooling that make software development, testing, and releases more efficient and reliable. The role requires cloud architecture knowledge, CI/CD experience, Terraform and Bazel proficiency, and software development skills in Go, Python, or C++.
Build and operate scalable control-plane and data-plane infrastructure for distributed AI workloads, including Ray cluster orchestration, scheduling, observability, and accelerator integration. Requires a bachelor's degree or equivalent experience, 3+ years of production coding, cloud-native expertise, Kubernetes, and Go/Python proficiency.
Leads infrastructure and platform strategy for a production healthcare AI platform, owning AWS, reliability, disaster recovery, compliance, CI/CD, and secure AI-agent operations. Requires deep cloud and Terraform expertise, audit-cycle experience, and prior technical leadership.
The Production Engineer will build and operate secure, scalable, and reliable infrastructure and production services, while developing engineering frameworks and supporting on-call operations. The role requires 5+ years of production, site reliability, or DevOps experience and familiarity with AWS, Kubernetes, and Terraform.
Own the reliability, resilience, observability, and automation of AWS and Kubernetes infrastructure supporting production products and AI/ML workloads. The role requires 4+ years of cloud infrastructure experience, strong Kubernetes and Terraform expertise, and senior-level incident response and software engineering skills.