CVS Health is committed to building a world of health around every individual. The Lead Director, AI Network & Data Center Engineering is responsible for leading the strategy, design, and delivery of network infrastructure supporting enterprise applications and AI/ML workloads.
Responsibilities:
- Define and enforce network architecture standards, security controls, and compliance policies
- Establish best practices for configuration, deployment, and operations
- Ensure alignment with enterprise governance frameworks (HIPAA, NIST, PCI, HITRUST, CSA)
- Advise executives on network strategy, roadmaps, emerging technologies, and industry trends
- Partner with cross-functional teams (platform, cloud, security, infrastructure) to deliver integrated, secure and scalable solutions
- Engage vendors on solution architectures, design, contracts, and delivery
- Design and implement:
- Data center fabrics (leaf-spine, Clos, EVPN/VXLAN)
- AI/GPU-optimized networking (low latency, high throughput) and related technologies
- Optimize performance for GPU workloads (RDMA, RoCE, ECN, PFC)
- GCP and multi-cloud network architectures
- Routing & switching (BGP, OSPF, MPLS, STP; Cisco, Juniper, Arista)
- High-speed networking (100/200/400Gb+)
- GPU cluster networking, traffic optimization, and congestion management
- Drive network monitoring and performance optimization (SolarWinds, NetFlow, Wireshark)
- Support SaaS and multi-tenant environments with zero trust patterns
- Oversee 24/7 network operations across cloud and data centers
- Lead incident response and ensure high availability, resiliency and scalability across network and cloud environments
- Build and mentor high-performing engineering teams
- Foster a culture of collaboration, accountability, and continuous learning
- Advance networking capabilities for:
- AI/ML infrastructure and GPU clusters (NVIDIA architectures)
- High-performance networking (RoCE, PFC, ECN)
- Champion automation and Infrastructure as Code (IaC)
- Contribute to industry and open-source initiatives
- Own network architecture strategy and roadmap
- Define standards for AI, cloud, and data center environments
- Support budgeting, risk management, and long-term planning
Requirements:
- 10+ years' experience in network and/or data center engineering
- 5+ years' experience in a strategic leadership position with people management
- 5+ years' experience in large scale enterprise environments
- Bachelor's degree or equivalent experience (High School Diploma and 4 years relevant experience)
- Strong expertise in large-scale data center design, security, and high availability
- Advanced problem-solving, execution, and stakeholder management skills
- Experience with AI factory / GPU cluster environments at scale
- Enterprise networking, Cisco Nexus/ACI or similar platforms
- Network automation / IaC, SDN, service mesh, and cloud-native security frameworks
- High availability/disaster recovery design and cyber resilience
- Certifications: CCIE/CCNP, PCNSE, CISSP, Cloud networking