SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Lambda is a leader in AI cloud infrastructure serving tens of thousands of customers, from AI researchers to enterprises and hyperscalers. The company's mission is to make compute as ubiquitous as electricity and give everyone access to superintelligence.
We are seeking a Senior Director of Network Software Services to lead one or more teams building the software that enables Lambda to run a secure, performant, and available AI Cloud at massive scale. This is a leader-of-leaders role reporting to the VP of Cloud and AI Networking.
Key Responsibilities:
- Own the networking services that deliver network monitoring, security, performance, and availability goals
- Own the architecture and operational excellence of large-scale distributed network systems, control planes, and data paths
- Own company-level goals including programmatically improving security posture, observability of network health, internet and backbone traffic engineering capability
- Set the multi-year technical strategy for network software services to scale ahead of demand
- Hire, develop, and retain engineers and the leaders who manage them; design team structure as the organization grows
- Manage through leads and senior individual contributors, providing regular 1:1s, clear performance feedback, growth planning, and sponsorship of meaningful work
- Partner deeply with network engineering, operations, capacity, and product teams to build services that improve Lambda connectivity quality
- Set the operational bar for your organization: SLOs, on-call health, incident response, and blameless postmortems
- Contribute to organization-wide engineering process improvements in planning, code shipping, and incident learning
- Partner with Principal Engineers and technical leadership to maintain high engineering standards across design, code quality, and operational reliability
- Represent your organization's work and technical direction to senior leadership, enterprise customers, and partners
Requirements:
- 8+ years building highly available, large-scale distributed systems for cloud infrastructure or enterprise network architecture
- 5+ years leading software engineering teams
- 2+ years managing managers
- Software engineering background with shipped production systems and credible engagement on architecture, technical tradeoffs, and code quality
- Deep expertise across networking protocols and data-center/cloud networking (BGP, EVPN/VXLAN, software-defined networking, traffic engineering)
- Deep expertise in zero-trust security, CI/CD automation, and cloud platforms such as AWS or OCI
- Proven track record of building and scaling engineering organizations that deliver mission-critical, high-performance cloud connectivity at scale
- Ability to foster a data-driven, automation-first, AI-enabled, high-velocity organization
- Ability to translate ambiguous business and product goals into clear team priorities and executable engineering plans
- Strong judgment about when to go deep technically, when to delegate, and when to escalate
- Track record of project and product delivery in fast-paced, high-pressure environments
- Excellent written and verbal communication skills; ability to use high-quality written artifacts to drive decisions and alignment
- BS or MS in Computer Science, Electrical Engineering, or related field, or equivalent practical experience
Nice to Have:
- Experience in GPU cloud, HPC, or AI/ML infrastructure environments
- Deep familiarity with large-scale data center or cloud networking (fabric build-out, capacity modeling, network supply chain)
- Familiarity with networking fundamentals and configuration management
- Prior experience at a high-growth infrastructure or cloud company