We are hiring a Staff Network Engineer to help build and operate the backbone of a carrier-grade, high-performance AI infrastructure. As a key technical contributor, you will design, deploy, and support large-scale network systems that connect our GPU clusters, high-throughput storage, and compute environments across geographically distributed data centers.
You will work alongside Principal Engineers and cross-functional teams to deliver automation-driven, low-latency networking designed for the scale and intensity of AI workloads and HPC environments.
Key Responsibilities
Implement and maintain high-throughput, low-latency networks supporting AI Factory workloads and distributed training infrastructure.
Work hands-on to deploy, configure, and troubleshoot routing, switching, optics, and interconnect systems across data centers.
Operate and optimize layer 2/3 network services: BGP, EVPN/VXLAN, OSPF, MPLS, QoS, and ACLs.
Work with Infiniband Networking Systems and Nvidia Fabric Manager (UFM)
Develop and maintain network automation (e.g., Ansible, Python, Terraform) for provisioning, compliance, and operational workflows.
Monitor network health and performance using telemetry tools and help scale observability platforms.
Participate in the incident response rotation and perform root cause analysis on service-impacting events.
Maintain configuration standards, documentation, and change management in line with infrastructure governance processes.
Collaborate with the Principal Network Engineer on architectural decisions and vendor evaluations.
Qualifications
Required:
5–8+ years of hands-on experience in large-scale network engineering, data center networks, or service provider infrastructure
Strong knowledge of IP networking, BGP, OSPF, EVPN/VXLAN, and L2/L3 design principles
Experience configuring and operating Arista, Juniper, or Cisco platforms in production environments
Proficiency in scripting or automation (e.g., Python, Bash, Ansible)
Solid troubleshooting skills and experience with real-time diagnostics and packet analysis
Familiarity with monitoring and telemetry tools (e.g., Prometheus, Grafana, sFlow, InfluxDB)
Preferred:
Experience in AI, HPC, or GPU-based infrastructure
Exposure to carrier-grade architectures, DCI, and optical transport systems
Exposure to Nvidia Infiniband Networking systems and components.
Understanding of network segmentation, security policies, and zero-trust principles
Comfortable working in 24/7 operational environments and on-call rotations
Voltage Park is an equal opportunity employer and makes employment decisions on the basis of merit. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status, or any other characteristic under federal, state, or local law. If you require an accommodation during the job application process, please notify your recruiter.
Compensation Range: $160K - $210K
#BI-Remote
Top Skills
Similar Jobs at Voltage Park
What you need to know about the Colorado Tech Scene
Key Facts About Colorado Tech
- Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
- Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
- Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
- Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
- Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute