DriveNets Logo

DriveNets

Group Lead - Infrastructure Services

Posted Yesterday
Be an Early Applicant
Remote
Hiring Remotely in United States
Expert/Leader
Remote
Hiring Remotely in United States
Expert/Leader
Leads the Infrastructure Services team of Solution Engineers and Solutions Architects delivering AI/HPC infrastructure solutions. Oversees customer engagements from pre-sales architecture and proof-of-concept execution through deployment, benchmarking, and operations. Acts as the senior technical escalation point, partners with Sales, Product, and Engineering, guides performance optimization, develops deployment processes, manages technology partnerships, creates technical content, and recruits and mentors team members.
The summary above was generated by AI
Description

Title: Group Lead - DriveNets Infrastructure Services (DIS)

#LI-Remote

Remote US - East Coast and Central Time zone preferred

About the Company 

DriveNets is a leader in high-scale networking software for AI infrastructure and service providers. The company pioneered a disaggregated networking architecture that transforms the economics of large-scale networks while maximizing performance, utilization, and operational efficiency. DriveNets-powered networks are deployed by global leaders, including AT&T and Comcast, supporting more than 30% of total U.S. internet traffic. DriveNets AI Fabric delivers full-stack networking for AI infrastructures, providing the highest-performance, Ethernet-based alternative to InfiniBand. The solution is deployed by hyperscalers, NeoClouds, and enterprises worldwide. With over $1B raised, DriveNets continues to push the boundaries of modern networking infrastructure.

The Role

DriveNets is seeking a Group Lead for its Infrastructure Services (DIS) team to be a key member of our customer-facing technical organization. Join a dynamic and forward-thinking company at the forefront of AI infrastructure. We leverage advanced technologies to develop innovative solutions that drive efficiency, scalability, and exceptional compute performance. Collaborate with the industry's best as we partner with hyperscalers, emerging NeoClouds, and enterprises building large-scale AI/HPC clusters, shaping the future of disaggregated AI networking and compute infrastructure. Our environment fosters creativity, teamwork, and growth, and offers you the opportunity to make a meaningful impact while leading a high-performing team on groundbreaking deployments.

As Group Lead for DIS, you will manage and develop a team of Solution Engineers and Solutions Architects responsible for designing, deploying, and optimizing DriveNets' AI/HPC infrastructure solutions at customer sites. You will provide technical leadership across the full customer lifecycle - from pre-sales architecture and POC execution through deployment, performance benchmarking, and ongoing operations. You will work cross-functionally with Sales, Product Management, and Engineering to ensure customer success, drive product feedback, and continuously raise the bar for technical delivery quality across the team.

Responsibilities

  • Lead and develop the DIS team - a group of Solution Engineers and Solutions Architects - setting technical direction, managing execution, and fostering a culture of ownership, learning, and customer focus.
  • Oversee end-to-end customer engagement for DIS - from pre-sales technical support and solution architecture through POC planning, deployment execution, and post-deployment operations.
  • Serve as the senior technical escalation point for customer infrastructure challenges, including AI cluster performance issues, networking design trade-offs, and operational reliability concerns.
  • Partner with Sales Account Managers to support business opportunities, lead technical responses to RFP/RFQs, and influence technical decision-makers at the VP and CxO level.
  • Guide the team in conducting performance benchmarking activities - including NCCL/RCCL, RDMA, and LLM benchmarks - and ensure results are translated into actionable product and deployment insights.
  • Work with Product Management and Engineering to funnel customer requirements, field observations, and performance data into the product roadmap and development backlog.
  • Define and drive internal processes for deployment planning, operational readiness, monitoring standards, and technical documentation across the DIS team.
  • Build and maintain relationships with compute, NIC, and storage partners to support joint POCs, reference deployments, and solution validation.
  • Represent DriveNets at industry events and conferences, and contribute to external technical content including white papers, blogs, and design guides.
  • Recruit, mentor, and grow team members, and establish clear performance goals aligned with business objectives.
Requirements

What we need to see:

  • 10+ years of experience in AI/HPC infrastructure, data center networking, or solutions architecture, with at least 2-3 years in a technical leadership or team lead capacity.
  • Hands-on technical depth across both compute infrastructure (GPU clusters, Linux systems, AI workloads) and data center networking (routing, switching, fabric design), with the ability to engage credibly across both disciplines.
  • Proven experience leading customer-facing technical teams through complex deployment and POC cycles in AI/HPC or data center environments.
  • Strong understanding of AI cluster architecture - including GPU platforms (NVIDIA, AMD), RDMA networking, storage connectivity, and the interaction between compute, network, and storage layers.
  • Experience with performance benchmarking methodologies (NCCL/RCCL, RDMA, LLM workloads) and the ability to interpret and act on results at a system level.
  • Demonstrated ability to work cross-functionally with Sales, Product Management, and Engineering teams, translating customer feedback into product improvements and go-to-market strategy.
  • Excellent communication and presentation skills, with proven ability to influence technical and executive stakeholders at customer organizations.
  • Ability to write extensive technical content (white papers, technical briefs, design guides, etc.) for external audiences with a balance of technical accuracy and clear messaging.
  • Ability to travel domestic and international.

Ways to stand out from the crowd:

  • Deep familiarity with AI-relevant infrastructure technologies - InfiniBand, RoCEv2, lossless Ethernet (PFC, ECN), GPU, NIC, DPU, and accelerated computing platforms.
  • Hands-on experience deploying and operating large-scale AI/HPC clusters, including GPU resource scheduling (Slurm, Kubernetes), monitoring (Prometheus, Grafana, DCGM), and operational tooling.
  • Understanding of scale-up (NVLink, UALink) and scale-out (Enhanced Ethernet, UEC, InfiniBand) interconnect technologies and their design trade-offs.
  • Experience with CCL tuning (NCCL/RCCL), GPU environment setup, and performance optimization across large multi-node GPU clusters.
  • Familiarity with AI/ML frameworks (PyTorch, TensorFlow) and how workload characteristics interact with infrastructure design decisions.
  • Proven experience with one or more Tier-1 Clouds (AWS, Azure, GCP, or OCI) or emerging NeoClouds, and cloud-native architectures and software.
  • Background in data center operations fundamentals - networking, cooling, power, and rack-level design.
  • Experience engaging compute, NIC, or storage vendors on joint solution definition, reference architecture development, or benchmarking programs.

EDUCATION

BS/MS/PhD in Electrical/Computer Engineering, Computer Science, Physics, or other Engineering fields, or equivalent experience.

More About DriveNets  

Based in Israel with locations in Romania, US, India and Japan as well as extended teams, DriveNets operations cover more than twelve countries. With recognition by industry analysts and through partnerships with market leaders such as AMD, Broadcom, Dell and others, DriveNets is pushing market momentum, delivering the scale and efficiency that modern AI workloads demand. Visit our website:  https://drivenets.com/company

 If your experience is close but doesn’t fulfil all requirements, please submit your application. DriveNets is on a mission to build a special company comprised of individuals with different backgrounds, perspectives, and experiences. 

 DriveNets is an equal opportunity employer. We do not discriminate based on upon race, religion, national origin, sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with disability, or other applicable legally protected characteristics. 

  

Similar Jobs

48 Minutes Ago
Easy Apply
Remote or Hybrid
Easy Apply
170K-190K Annually
Senior level
170K-190K Annually
Senior level
Fintech • Information Technology • Software • Financial Services
Design and build Alloy’s go-to-market technology architecture, including CRM integrations, iPaaS automations, data pipelines, and workflow improvements. Develop SQL, Python, and JavaScript scripts for data manipulation and automation; monitor data quality and system performance; support reporting and BI dashboards; troubleshoot integration issues; and advise on GTM technology, vendors, and build-versus-buy decisions. Partner with RevOps, Sales, Marketing, Customer Success, IT, and leadership to translate business needs into scalable technical solutions.
Top Skills: APIsChurnzeroDbtGainsightGongHookJavaScriptLookerN8NOutreachPythonSalesforceSalesloftSQLTableauTray.IoWorkatoZapier
49 Minutes Ago
Remote or Hybrid
85K-105K Annually
Senior level
85K-105K Annually
Senior level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Supports NBCUniversal’s ad sales systems by owning tickets, troubleshooting software, integrations, calculations, and backend data issues. Monitors systems, manages defects, documents procedures, coordinates with vendors and engineering teams, and recommends process improvements and automation. The role provides stakeholder communication, incident management, alerting and logging guidance, and occasional evening and rotating weekend on-call coverage during special events.
Top Skills: Ai ToolsAPIsData StreamsMicroservicesRobotic Process Automation (Rpa)ScriptingSQLTest Automation
49 Minutes Ago
Remote or Hybrid
165K-210K Annually
Senior level
165K-210K Annually
Senior level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Leads automation engineering for NBCUniversal’s data collaboration platform. Designs AI agents, RAG workflows, reusable tools, APIs, and production-grade Python services supporting audience activation, measurement, and reporting. Establishes agent evaluation, monitoring, reliability, testing, and safety standards. Oversees clean room libraries and scalable data workflows while partnering across product, engineering, operations, and data platform teams. Hires, mentors, and develops software and AI engineers and drives technical direction, maintainability, and operational excellence.
Top Skills: Ci/CdDatabricksDatabricks Clean RoomsHabuLangchainLanggraphLiverampLlmsPythonRetrieval-Augmented Generation (Rag)SnowflakeSnowflake Clean RoomsSnowflake CortexVector Databases

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account