Skip to content

Senior Network Engineer

Senior Network Engineer responsible for designing, implementing, and maintaining high-performance compute network infrastructure for AI systems. Requires 8+ years experience with large-scale data center networks, deep expertise in routing/switching protocols, automation, and multi-vendor hardware.

About the job

Responsibilities

  • Design, deploy, manage and maintain global multi-vendor, multi-protocol high performance compute networks.
  • Analyze data to diagnose and identify root causes to network issues to minimize downtime.
  • Evaluate and recommend network technologies, hardware, and software solutions.
  • Participate in design reviews to ensure the proposed network architecture aligns with business needs and is optimized for performance, scalability, and reliability.
  • Manage relationships with external vendors and partners to test and verify hardware and software selections.
  • Develop and deploy systems and tools to keep all networks running reliably and efficiently.
  • Establish and implement industry best practices and contribute to the design of new scalable network solutions.
  • Ensure compliance with IT governance standards and best practices.
  • Lead projects to address complex technical challenges, directly contributing to roadmaps and partner alongside the best engineers in the industry to develop world-class solutions.

Requirements

  • 8+ years of professional experience building, managing, and supporting large-scale hybrid data center networks (excluding enterprise networks).
  • High level of proficiency with TCP/IP networking architecture and technologies such as BGP, OSPF, VXLAN, EVPN, and QoS.
  • Experience developing network automation pipelines using Python, Ansible, or other languages/tools utilized in infrastructure automation.
  • Proficient in using tools such as Wireshark, tcpdump, nmap, MTR, and curl to identify connectivity issues, latency problems, and network bottlenecks.
  • Experience designing and supporting multi-tenant networks.
  • Hands-on experience deploying and supporting network devices from Cisco, Arista, Juniper, and Mellanox.
  • Experience working with cloud networks such as AWS, GCP, and Azure.
  • Solid experience working in and troubleshooting within a Linux environment.

Preferred

  • Knowledge of RoCE and Infiniband protocols a plus.
  • Experience with Docker, Kubernetes, or Slurm a plus.
  • Understanding of AI training workloads and the demands they exert on networks a plus.

Compensation

The US base salary range for this full-time position is: $190,000 - $270,000 + equity + benefits.

Skills

BGP, Ospf, Vxlan, Evpn, Qos, Python, Ansible, Wireshark, Tcpdump, Nmap, Mtr, Cisco, Arista, Juniper, Mellanox

Astra

Astra

United States

Senior Platform Engineer
$190k+/yrRemote5+ YOEDevOps / SRE

Build and operate core platform infrastructure, developer tooling, CI/CD, observability, and cloud reliability systems for a regulated payments platform. Requires 5+ years of infrastructure or backend experience, strong infrastructure-as-code skills, and production cloud expertise.

Applied Intuition

Applied Intuition

Sunnyvale, CA

Senior Software Engineer - Cloud Infrastructure
$190k+/yrOn-site5+ YOEDevOps / SRE

Build and operate multi-cloud, multi-cluster infrastructure and platform primitives for large-scale simulations and enterprise AI workloads. The role requires 5+ years in infrastructure, platform, SRE, or DevOps systems, strong Kubernetes and cloud expertise, production programming skills, and Infrastructure as Code experience.

inKind

inKind

Austin, TX

Senior Platform Engineer
$190k+/yrRemote8+ YOEDevOps / SRE

Own and evolve AWS cloud infrastructure, deployment, reliability, observability, and security for a growing financial and hospitality technology platform. The hands-on role requires 8+ years operating production cloud infrastructure, strong AWS and container orchestration expertise, and experience with migrations and incident response.

Mercury

Mercury

San Francisco, CA
Senior Software Engineer - SRE
$190k+/yrRemote5+ YOEDevOps / SRE

Senior SRE who embeds with product teams to improve reliability, observability, performance, and incident preparedness. The role requires SRE or DevOps experience, strong PostgreSQL and Temporal expertise, and familiarity with observability platforms and OpenTelemetry.

Idme

Idme

McLean, VA
Senior Software Engineer – Platform & Data Infrastructure
$191k+/yrOn-site8+ YOEDevOps / SRE

Senior engineer responsible for scaling and operating multi-region Kubernetes, GitOps, Infrastructure as Code, security governance, and data-platform infrastructure. The role requires 8+ years of platform, SRE, or cloud data infrastructure experience and strong Kubernetes and Terraform expertise.