senior software engineering Site Reliability Engineer ic · Posted Jan 15, 2026
$136,600 – $184,800
USD per year

About this role

Amazon is hiring a senior-level Site Reliability Engineer in the software engineering function based in Herndon, VA. The posting calls out experience with AWS, Networking, Machine Learning, Cloud Computing. Compensation is listed at $136,600–$184,800 per year.

Role
Site Reliability Engineer
Function
software engineering
Level
senior
Track
Individual contributor
Employment
Full-time
Location
Herndon, VA
Department
Systems, Quality, & Security Engineering
Posted
Jan 15, 2026
AI Summary
Drive reliability risk identification, assessment, and mitigation for AWS datacenter infrastructure equipment. Conduct root cause analysis of critical failures and implement continuous improvements. Requires expertise in physics-of-failure approaches, statistical analysis, vendor management, and system-level reliability engineering.

More roles at Amazon

Manufacturing Engineer, Amazon Leo
Redmond, WA · mid
System Design
Manufacturing Engineer, Amazon Leo
Redmond, WA · mid
System Design
Sr. FPGA Engineer, Amazon Leo
Redmond, WA · senior
Python Git CI/CD
Senior Customer Success Manager, Strategic Account Services (LPSAS)
Virtual, Costa Rica · senior
Salesforce Agile Data Analytics
Facilities Coordinator III, NA AMOC
San Jose, Costa Rica · mid
Incident Response
All Amazon jobs →

Job description

from Amazon careers

As an Infrastructure Reliability Engineer you will be proactively driving the reliability risk identification, assessment and mitigation for datacenter infrastructure equipment (Example: Air Handling Units, LV Generator, MV Transformers, LV SWGR, Breakers, UPS, Chillers etc.). You will also be responsible for root cause analysis of critical equipment failures and drive the continuous improvements to improve datacenter availability for AWS customers. You will work closely with both internal and outside partners including suppliers to drive key aspects of product specification, risk identification plan and execution. You must be ownership minded, independent, action and results oriented to succeed in an open collaborative environment. The candidate should have experience in using Physics-of-Failure based approach to develop and implement both analytical and empirical approaches for product quality/reliability risk identification and assessment during product design, manufacture as well as deployment stages. The individual should be able to drive AWS application-specific requirements in carrying out both lifecycle environmental and operational stress driven risk analysis, including thermal, electrical, chemical and mechanical stresses so to identify overstress and fatigue-related product weaknesses. Candidate should be capable of evaluating not only product design quality/reliability risks, but also have the skills and experiences in assessing electronics manufacture process related quality/reliability issues.…

This is an excerpt. Read the full job description on Amazon careers →
All software engineering jobs software engineering in Herndon, VA Jobs in Herndon, VA software engineering salaries software engineering career path
All Amazon Jobs Browse software engineering roles senior positions