Network Engineer, Deployment & Integration
Hands-on network engineer deploying and validating large-scale AI datacenter fabrics, configuring switches, troubleshooting physical/optical layers, and coordinating cross-functional teams. Requires 3-7 years datacenter experience and 70-80% travel to onsite locations.
About the job
Responsibilities
Deployment Execution
- Deploy and validate datacenter network infrastructure including front-end fabric, back-end fabric, BMS, and management networks.
- Configure switches, install and validate optics, coordinate fiber/cabing work, and drive deployments through completion.
Physical Layer Validation
- Ensure physical connectivity meets production standards.
- Coordinate with structured cabling teams on fiber remediation, validate insertion loss and OTDR traces, troubleshoot optical layer issues, and document physical infrastructure as-builts.
Hardware Lifecycle Management
- Manage hardware logistics including device staging, rack/stack coordination, RMA processes, and DCIM updates.
- Track hardware inventory, coordinate vendor shipments, and ensure devices are ready when deployments need them.
Cross-Functional Coordination
- Partner with DC Operations, ICT teams, Hardware teams, and Network Engineering to drive deployments forward.
Documentation & Process Improvement
- Maintain accurate documentation of deployment activities including cutsheets, as-builts, validation results, and lessons learned.
- Identify gaps in deployment procedures and propose improvements.
Operational Support
- Provide backup operational support during and after deployments.
- Respond to incidents, execute troubleshooting procedures, and coordinate break-fix activities.
Requirements
- Datacenter Networking Foundation: 3-7 years in network engineering with hands-on datacenter experience. Understand modern datacenter fabrics (EVPN/VXLAN, BGP, CLOS architectures). Comfortable with CLI, configuration management, and network validation.
- Hands-On Execution Mindset: Thrive in field environments, comfortable pulling cable, configuring switches, troubleshooting optical layer issues.
- Strong Troubleshooting Skills: Diagnose issues across physical and logical layers. Read OTDR traces, validate insertion loss, debug BGP sessions.
- Coordination & Communication: Clear communicator across teams, document work clearly.
- Self-Directed Learning: Learn quickly, take ownership of ramping up.
- Travel Ready: Comfortable with 70-80% travel to onsite deployments.
Nice to Haves
- AI Fabric Experience: RDMA (RoCEv2), lossless Ethernet (PFC, ECN), high-performance compute fabrics.
- Vendor Platform Knowledge: Arista, Juniper, NVIDIA networking platforms.
- Physical Layer Expertise: Structured cabling standards, fiber optics (SMF/MMF), insertion loss budgets, optical validation tools.
- Automation Exposure: Network automation, configuration templating, scripting (Python, Ansible).
- DCIM/Asset Management: Datacenter infrastructure management tools, asset tracking.
Compensation
- Base salary range: $150,000 - $250,000 per year, depending on experience, skills, qualifications, and location.
- Competitive total compensation package (salary + equity).
- Retirement or pension plan, health/dental/vision insurance, generous PTO.
Skills
Evpn/Vxlan, BGP, Clos, Arista, Juniper, Nvidia, Rocev2, Rdma, Pfc, Ecn, Otdr, Python, Ansible, Dcim
Similar jobs
DevOps / SRE jobsBuild and operate a highly available, multi-region PostgreSQL platform, developing automation, monitoring, disaster recovery, and performance tooling. Requires experience with large-scale PostgreSQL clusters, infrastructure as code, scripting, containers, and observability.
Leads on-site deployment of data center physical infrastructure, managing contractors, performing QA/QC on fiber optics and cabling, and ensuring compliance with standards. Requires 5+ years experience, SME-level fiber optic expertise, bachelor's degree, and 40% travel readiness.
The Python Engineer will improve and operate trading systems, support integrations with asset classes and prime brokers, and handle monitoring, incidents, and performance optimization. The role requires 3+ years of experience, strong Python and Linux skills, and familiarity with market data and order-entry systems.
Build IT workflow automation and security tooling using Go, Temporal, Kubernetes, Terraform, and shell scripting. The role also supports internal IT systems and integrations across endpoint management, identity, access, and security administration.
Designs and operates foundational developer-infrastructure services for CI, builds, deployments, and testing. The role requires senior-level systems engineering, end-to-end service ownership, and cross-functional technical leadership.