Network Development Manager, Capacity Restoration Team
Amazon Web Services (AWS)
Date: 2 weeks ago
City: Hyderabad, Telangana
Contract type: Full time
Description
AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain — and we’re looking for talented people who want to help.
You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.
The Capacity Restoration Team (CRT) is responsible for restoring backbone and border network capacity impacted by failures, maintenance events, and service disruptions across Amazon's global network infrastructure. As an Network Development Manager for CRT, you will lead a team of Network Development Engineers focused on building and operating capacity restoration solutions, automation tooling, and operational processes that ensure rapid recovery of out-of-service network capacity.
This is a front-line management role where you will be hands-on with your team, driving day-to-day execution while contributing to the team's technical direction. You will work within a defined strategic framework, owning the delivery of capacity restoration products and driving engineering excellence across your team. Your work directly impacts network availability, Mean Time to Remediate (MTTR), and capacity utilization across Amazon's backbone and border fabric.
Key job responsibilities
Team Leadership & People Management
The Capacity Restoration Team (CRT) operates within BERE Operations and is responsible for ensuring Amazon's backbone and border network capacity is rapidly restored following failures, fiber cuts, maintenance events, and service disruptions. The team works at the intersection of network engineering, automation, and operational intelligence to minimize time-to-restore and prevent customer-impacting capacity shortfalls.
CRT is currently being established across Seattle and Hyderabad with a mission to systematically reduce out-of-service capacity and improve restoration rates year over year. The team collaborates closely with backbone topology, border engineering, optical networking, and network controller teams to deliver holistic capacity restoration outcomes across Amazon's global network infrastructure.
Basic Qualifications
Company - ADSIPL - Telangana
Job ID: A10461577
AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain — and we’re looking for talented people who want to help.
You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.
The Capacity Restoration Team (CRT) is responsible for restoring backbone and border network capacity impacted by failures, maintenance events, and service disruptions across Amazon's global network infrastructure. As an Network Development Manager for CRT, you will lead a team of Network Development Engineers focused on building and operating capacity restoration solutions, automation tooling, and operational processes that ensure rapid recovery of out-of-service network capacity.
This is a front-line management role where you will be hands-on with your team, driving day-to-day execution while contributing to the team's technical direction. You will work within a defined strategic framework, owning the delivery of capacity restoration products and driving engineering excellence across your team. Your work directly impacts network availability, Mean Time to Remediate (MTTR), and capacity utilization across Amazon's backbone and border fabric.
Key job responsibilities
Team Leadership & People Management
- Manage a team of 8-12 Network Development Engineers (NDEs) and Network Engineers (NEs) focused on capacity restoration execution, automation development, and operational support
- Hire, develop, coach, and promote engineers; conduct effective performance reviews and create individualized growth plans
- Build a high-performing team culture with emphasis on operational rigor, ownership, and continuous learning
- Establish and maintain a sustainable on-call rotation; ensure incident response coverage across time zones
- Manage day-to-day team operations including sprint planning, backlog prioritization, and resource allocation
- Foster collaboration between your India-based team and the Seattle CRT counterpart
- Own delivery of capacity restoration products and features within your team's scope (e.g., automated restoration workflows, capacity dashboards, Ariadne controller integration features)
- Drive hands-on technical decisions at the component and system level; review designs, code, and architecture proposals from your team
- Translate team-level strategy into actionable engineering plans with clear milestones and deliverables
- Identify and resolve technical blockers; make trade-offs between quality, speed, and scope with awareness of long-term implications
- Drive automation of manual capacity restoration processes; measure and improve automation coverage
- Ensure engineering solutions are scalable, maintainable, and aligned with broader network architecture patterns
- Own capacity restoration SLAs for your team's scope; drive Mean Time to Remediate (MTTR) improvements
- Define and track operational metrics including restoration success rate, automation levels, and escalation percentages
- Drive root-cause analysis and Corrections of Error (COE) processes; ensure systemic fixes are implemented
- Manage technical debt and deferred maintenance; allocate engineering time to reduce operational risk
- Ensure team compliance with Amazon security, data handling, and operational policies
- Participate in and drive operational reviews (WBR/MBR); present team metrics and improvement plans
- Partner with backbone operations, border engineering, optical networking, and network controller teams on shared restoration outcomes
- Represent your team's work to peer managers and senior leadership; communicate progress, risks, and dependencies clearly
- Collaborate with the CRT Seattle team to ensure consistent processes, tooling, and follow-the-sun operational coverage
- Work with recruiting to build hiring pipelines; actively participate in interview loops and hiring decisions.
The Capacity Restoration Team (CRT) operates within BERE Operations and is responsible for ensuring Amazon's backbone and border network capacity is rapidly restored following failures, fiber cuts, maintenance events, and service disruptions. The team works at the intersection of network engineering, automation, and operational intelligence to minimize time-to-restore and prevent customer-impacting capacity shortfalls.
CRT is currently being established across Seattle and Hyderabad with a mission to systematically reduce out-of-service capacity and improve restoration rates year over year. The team collaborates closely with backbone topology, border engineering, optical networking, and network controller teams to deliver holistic capacity restoration outcomes across Amazon's global network infrastructure.
Basic Qualifications
- Knowledge of network design, protocols and troubleshooting
- Experience in managing and troublshooting network, or experience with automation and any version control tools and experience in technical support
- Experience communicating and presenting to senior leadership
- 4+ years of network development engineering experience with at least 2+ years in a people management role
- Demonstrated experience hiring, developing, and managing a team of engineers (6+ direct reports)
- Track record of delivering engineering projects on time with high quality in a fast-paced environment
- Ability to translate business requirements into technical execution plans with clear milestones
- Bachelor's degree in Computer Science, Electrical Engineering, Network Engineering, or equivalent experience
- Experience with capacity management systems, network controllers, or traffic engineering at scale
- Familiarity with backbone/border network failure modes, restoration workflows, and Out-of-Service (OOS) capacity management
- Experience with optical networking (DWDM/CWDM transport systems, optical module qualification, fiber optic infrastructure)
- Background in large-scale distributed network architectures (data center, backbone, or border fabrics)
- Experience driving operational metrics (SLOs, MTTR, capacity utilization) and using data to prioritize engineering work
- Familiarity with Python, network automation frameworks, or infrastructure-as-code tools
- AWS/cloud networking knowledge or experience with network modeling and simulation tools
- Master's degree in a related technical field
Company - ADSIPL - Telangana
Job ID: A10461577
How to apply
To apply for this job you need to authorize on our website. If you don't have an account yet, please register.
Post a resumeSimilar jobs
Staff DevOps Engineer
Velotic Software,
Hyderabad, Telangana
2 weeks ago
Velotic is a leading independent industrial software company providing data‑driven solutions that improve manufacturing efficiency, productivity, and operational insight. Serving customers across manufacturing, oil & gas, utilities, and infrastructure, the company generates more than $300 million in revenue. Velotic’s portfolio, anchored by Proficy, Kepware, and ThingWorx, supports the growing data and performance requirements of industrial operators globally, with an emphasis...
Lead Engineer
Johnson & Johnson Innovative Medicine,
Hyderabad, Telangana
2 weeks ago
At Johnson & Johnson, we believe health is everything. Our strength in healthcare innovation empowers us to build a world where complex diseases are prevented, treated, and cured, where treatments are smarter and less invasive, and solutions are personal. Through our expertise in Innovative Medicine and MedTech, we are uniquely positioned to innovate across the full spectrum of healthcare solutions...
Workplace & Facilities Coordinator
DoorDash,
Hyderabad, Telangana
2 weeks ago
About the TeamYou'll be joining the Real Estate and Workplace Operations (REWO) team within DoorDash's Finance Organization. We design, build, and operate inclusive, inspiring, and functional spaces that help employees do their best work, removing physical friction so teams can focus on what matters.About the RoleThis role supports day-to-day Workplace and Facilities operations for our India site, helping ensure a...