Director, Network Capacity Automation
Lambda Labs
- Location
- Bellevue Office
- Employment
- Full Time
- Work model
- Remote
- Level
- Staff
- Posted
- 1h ago
Skills
About this role
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. *Note: This position requires presence in our Bellevue or San Francisco office location 4 days per week; Lambda’s designated work from home day is currently Tuesday. Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. Our vision is bold and is not an incremental exercise. We will continually re-evaluate and reinvent our current ways of working and the automation underneath all of it, while operating the existing network flawlessly for all customer workloads. Lambda has 10x'd over the last three years, and the network engineering organization is scaling to match. We're looking for a Director of Network Capacity Automation to lead one or more teams building the software that enables us to plan and deliver network capacity ahead of customer demand. This capacity is critical to ensure a flawless customer connectivity experience to our GPU cloud infrastructure. We're looking for a Director to own that outcome end to end — and to build the organization that delivers it. You'll hire and grow a talented team of software and network proficient engineers, shape team culture, and partner with technical leadership to deliver the mechanisms needed to stay ahead of our scale ambitions. This role will report to the VP of Cloud and AI Networking, working alongside a highly capable and passionate team of engineers. If you'd like to join a great team of customer obsessed engineers building the world's best AI cloud, come join us.
What You'll Do
Own network capacity end to end — planning, delivery, turn-up and lifecycle — and be accountable for capacity landing ahead of demand Build and lead the organization that delivers it: hire, develop and retain engineers and the leaders who manage them, and design the team structure as the org grows Set the multi-year technical strategy for capacity automation to ensure it continually scales Convert manual process into durable software — the goal is a capacity pipeline that runs without heroics, not a better-organised set of runbooks Own the operational bar for your org: how capacity work is planned, reviewed, shipped, measured and learned from when it goes wrong Manage through leads and senior ICs, developing both — regular 1:1s, clear performance feedback, growth planning, and sponsorship of meaningful work Partner and support our Interconnect team who negotiate, acquire and coordinate on external connectivity Partner with Principal Engineers and technical leadership to maintain a high engineering bar across design, code quality and operational reliability Partner with Infrastructure, Supply Chain, Data Center Operations, Product and Finance to align capacity roadmaps, forecast demand and manage long-lead dependencies Represent your org's work and technical direction to senior leadership, and to enterprise customers and partners where appropriate Contribute to org-wide engineering process improvements — how we plan, how we ship, how we learn from incidents You Have 8+ years of engineering management experience, including directly managing senior individual contributors Proven track record of building and scaling engineering organizations that deliver mission-critical, high-performance cloud connectivity serving millions of users Have a software engineering background — you've shipped production systems and can engage credibly on architecture,