Software Development Engineer, EC2 UltraServer Availability
Seattle, WA, United States • Posted June 02, 2026
Job Type:
Full-time
Location:
Seattle, WA
Posted:
June 02, 2026
Category:
other-general
Application Deadline:
June 08, 2026
Role Description
Description
The Software Development Engineer II will design, build, and maintain cloud-based repair and recovery workflows for NVIDIA GB200 / GB300 UltraServers, orchestrating repair and recovery operations from impairment detection through completed recovery. This role requires expertise in AWS services, system architecture, and cross-functional collaboration with Capacity Management, Hardware Engineering, and Datacenter Operations to manage AI/ML infrastructure.
Key job responsibilities
The Software Development Engineer (SDE II) on the EC2 UltraServer Availability team is responsible for ensuring high availability of customer GB200 and GB300 UltraServers by orchestrating complex repair and recovery workflows. Following are the core responsibilities
System Design & Architecture
* Design and architect solutions that are cross-functional to Capacity Management, Hardware Engineering, and Datacenter Operations
* Work in environments where the technolog...
The Software Development Engineer II will design, build, and maintain cloud-based repair and recovery workflows for NVIDIA GB200 / GB300 UltraServers, orchestrating repair and recovery operations from impairment detection through completed recovery. This role requires expertise in AWS services, system architecture, and cross-functional collaboration with Capacity Management, Hardware Engineering, and Datacenter Operations to manage AI/ML infrastructure.
Key job responsibilities
The Software Development Engineer (SDE II) on the EC2 UltraServer Availability team is responsible for ensuring high availability of customer GB200 and GB300 UltraServers by orchestrating complex repair and recovery workflows. Following are the core responsibilities
System Design & Architecture
* Design and architect solutions that are cross-functional to Capacity Management, Hardware Engineering, and Datacenter Operations
* Work in environments where the technolog...
Interested in this role?
Click the button below to start your application for Software Development Engineer, EC2 UltraServer Availability at Amazon.
Apply Now