Sr. Hardware Reliability Engineer, Infrastructure Reliability & Quality
Amazon · Herndon, VA · Systems, Quality, & Security Engineering
About this role
Amazon is hiring a senior-level Site Reliability Engineer in the software engineering function based in Herndon, VA. The posting calls out experience with AWS, Networking, Machine Learning, Cloud Computing. Compensation is listed at $136,600–$184,800 per year.
- Role
- Site Reliability Engineer
- Function
- software engineering
- Level
- senior
- Track
- Individual contributor
- Employment
- Full-time
- Location
- Herndon, VA
- Department
- Systems, Quality, & Security Engineering
- Posted
- Jan 15, 2026
More roles at Amazon
Job description
from Amazon careersAs an Infrastructure Reliability Engineer you will be proactively driving the reliability risk identification, assessment and mitigation for datacenter infrastructure equipment (Example: Air Handling Units, LV Generator, MV Transformers, LV SWGR, Breakers, UPS, Chillers etc.). You will also be responsible for root cause analysis of critical equipment failures and drive the continuous improvements to improve datacenter availability for AWS customers. You will work closely with both internal and outside partners including suppliers to drive key aspects of product specification, risk identification plan and execution. You must be ownership minded, independent, action and results oriented to succeed in an open collaborative environment. The candidate should have experience in using Physics-of-Failure based approach to develop and implement both analytical and empirical approaches for product quality/reliability risk identification and assessment during product design, manufacture as well as deployment stages. The individual should be able to drive AWS application-specific requirements in carrying out both lifecycle environmental and operational stress driven risk analysis, including thermal, electrical, chemical and mechanical stresses so to identify overstress and fatigue-related product weaknesses. Candidate should be capable of evaluating not only product design quality/reliability risks, but also have the skills and experiences in assessing electronics manufacture process related quality/reliability issues.…