Jobs and Careers
AM
Senior Hardware Dev Engineer, HWEng - Accelerated Server Platform Development
Amazon.comUnited Statesfull_timeVerifiedPosted 6 Mar 2023
💰 $213,600/yr($114,300/yr – $213,600/yr)
About the role
Do you enjoy designing scalable complex computing systems & solving difficult problems while driving influential change to a hyperscale environment? Are you curious about the systems used to run the largest cloud computing infrastructures in the world? Do you thrive in a fast paced and ever-changing environment?
Our team designs, builds and operates Amazon's fleet of complex computing systems (X, EC2 P, G, TRN, INF + more instance types). We solve systemic hardware issues and we build hardware and software systems to detect and mitigate future recurrences so that our our customers can experience the highest quality of service possible!
You will be responsible for owning the design and operations of a brand new segment of servers for the AWS fleet. As end to end owners of the complex server fleet, our team works closely with partners to root cause failures and drive changes back into our current & future designs. Nothing is complete without closed loop corrective actions which drive changes back into our development processes and behavior specifications.
As a member of the AWS Hardware Engineering organization, you will apply your technical experience and work with other subject matter experts in core component development, compute server development, networking development, custom hypervisor/virtualization development and other teams. You will be responsible for hardware and systems that improve how we detect, root cause, and remediate issues. You will lead cross functional investigations and define changes needed to deliver results and will have direct exposure to internal and external AWS customers. Ideal candidates will have a background in server development, system design, root cause, scoping complex issues, qualification, problem solving and developing corrective actions.
Key job responsibilities
As a member of the Enterprise, Trusted Compute & Accelerated Server Hardware Engineering team you will own and lead the design, development and root cause of a new segment of accelerated servers.
You will work closely with our customers to understand their technical needs and business goals, leveraging your experience with server design and the knowledge of various teams to architect the solutions that we will deploy at scale.
To deliver your products you will work with an interdisciplinary team of component, firmware, test, qualification, and integration engineers, and lead our design and manufacturing partners to bring these servers to the data center. After launch you will oversee the fleet of servers you develop, monitoring their quality and how they are meeting the customer requirements.
A day in the life
Your day to day responsibilities will include interfacing with our internal and external customers to understand project requirements and facilitate system development ontop of your server design. You will be responsible for learning operational challenges to our existing fleet with the goal of improving the current customer experience as well as developing improved systems for future designs. You will work directly with vendors and ODM/JDM design teams to develop and manufacture your product at scale.
About the team
The team is comprise of both Hardware Design Engineers, System Design Engineers, SDE's and Technical Program Managers, all with the common goal of delivering the best Accelerated Server fleet possible to our customers.
BS degree in EE, CE, or CS or relevant industry experience
5+ years of relevant work experience with server compute/storage platforms and debugging issues in a production environment.
3+ years of advanced support experience managing customer escalations to quickly identify root cause and resolve issues.
3+ years of Development experience in hardware/ firmware.
Demonstrated ability to engage technology developers in the industry to track and apply emerging technologies to innovative server designs
Working with interdisciplinary teams to execute product design from concept to production.
Developing functional specifications, design verification plans and functional test procedures
Commitment to rigorous testing practices and design process flows
Strong customer focus
Desire to work in a fast-paced environment
Ability to resolve complex issues in creative, efficient, and effective ways
Technical background with Memory, Storage, and PCIe debugging.
Technical background in server design or architecture
Developed Monitoring and alerting systems to quickly identify and categorize failures.
Leading cross functional teams through investigation and corrective actions.
BIOS/BMC/FW Debugging
Our team designs, builds and operates Amazon's fleet of complex computing systems (X, EC2 P, G, TRN, INF + more instance types). We solve systemic hardware issues and we build hardware and software systems to detect and mitigate future recurrences so that our our customers can experience the highest quality of service possible!
You will be responsible for owning the design and operations of a brand new segment of servers for the AWS fleet. As end to end owners of the complex server fleet, our team works closely with partners to root cause failures and drive changes back into our current & future designs. Nothing is complete without closed loop corrective actions which drive changes back into our development processes and behavior specifications.
As a member of the AWS Hardware Engineering organization, you will apply your technical experience and work with other subject matter experts in core component development, compute server development, networking development, custom hypervisor/virtualization development and other teams. You will be responsible for hardware and systems that improve how we detect, root cause, and remediate issues. You will lead cross functional investigations and define changes needed to deliver results and will have direct exposure to internal and external AWS customers. Ideal candidates will have a background in server development, system design, root cause, scoping complex issues, qualification, problem solving and developing corrective actions.
Key job responsibilities
As a member of the Enterprise, Trusted Compute & Accelerated Server Hardware Engineering team you will own and lead the design, development and root cause of a new segment of accelerated servers.
You will work closely with our customers to understand their technical needs and business goals, leveraging your experience with server design and the knowledge of various teams to architect the solutions that we will deploy at scale.
To deliver your products you will work with an interdisciplinary team of component, firmware, test, qualification, and integration engineers, and lead our design and manufacturing partners to bring these servers to the data center. After launch you will oversee the fleet of servers you develop, monitoring their quality and how they are meeting the customer requirements.
A day in the life
Your day to day responsibilities will include interfacing with our internal and external customers to understand project requirements and facilitate system development ontop of your server design. You will be responsible for learning operational challenges to our existing fleet with the goal of improving the current customer experience as well as developing improved systems for future designs. You will work directly with vendors and ODM/JDM design teams to develop and manufacture your product at scale.
About the team
The team is comprise of both Hardware Design Engineers, System Design Engineers, SDE's and Technical Program Managers, all with the common goal of delivering the best Accelerated Server fleet possible to our customers.
Basic Qualifications
BS degree in EE, CE, or CS or relevant industry experience
5+ years of relevant work experience with server compute/storage platforms and debugging issues in a production environment.
3+ years of advanced support experience managing customer escalations to quickly identify root cause and resolve issues.
3+ years of Development experience in hardware/ firmware.
Demonstrated ability to engage technology developers in the industry to track and apply emerging technologies to innovative server designs
Working with interdisciplinary teams to execute product design from concept to production.
Developing functional specifications, design verification plans and functional test procedures
Commitment to rigorous testing practices and design process flows
Strong customer focus
Desire to work in a fast-paced environment
Ability to resolve complex issues in creative, efficient, and effective ways
Preferred Qualifications
Masters/PhD Degree with 10+ years of relevant experience to Root Cause hardware issues on Server/Storage systems.Technical background with Memory, Storage, and PCIe debugging.
Technical background in server design or architecture
Developed Monitoring and alerting systems to quickly identify and categorize failures.
Leading cross functional teams through investigation and corrective actions.
BIOS/BMC/FW Debugging
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s