Are you passionate about building resilient, scalable systems that power critical services? As a System Development Engineer at Amazon, you will design and deliver technology solutions that improve operational health, reduce complexity, and enable teams to move faster with higher quality. You will own the engineering and operational excellence of your team's systems, identifying risks and eliminating them with permanent fixes. Your work will directly improve the customer experience by making systems more reliable, maintainable, and cost-effective. If you are a curious, self-starter who thrives on solving difficult problems and coaching others to do the same, this role offers the opportunity to make meaningful impact across teams.
This position requires that the candidate selected must currently possess and maintain an active TS/SCI security clearance with polygraph. The position further requires the candidate to opt into a commensurate clearance for each government agency for which they perform AWS work.
Key job responsibilities
- Design and deliver technology solutions that solve difficult business problems, ensuring they are thoroughly tested, pragmatic, efficient, and cost-effective.
- Improve your team's operational health by participating in design reviews, operational readiness reviews, and post-incident analyses to identify risks to resilience, then deliver projects that mitigate those risks.
- Deeply diagnose problems across hardware, software, and operating environments, resolving contributing causes of performance, reliability, and availability issues with permanent solutions.
- Automate repetitive tasks, develop monitoring metrics, and identify points of failure to continuously improve system resiliency and reduce operational burden.
- Write clear, accurate documentation and contribute to code reviews, providing meaningful feedback to peers while coaching others on identifying and eliminating risk.
A day in the life
You will start your day reviewing system health dashboards and addressing any operational alerts that surfaced overnight. From there, you might dive into designing an automation solution that eliminates a recurring manual process, participate in a post-incident review to uncover root causes, or pair with teammates on a code review. You will collaborate with partner teams to align on system dependencies and share operational best practices.
About the team
Our team is focused on building and maintaining systems that are reliable, scalable, and easy to operate. We invest in engineering excellence and automation so that we can deliver products and services to customers with less effort and better quality. We value collaboration, inclusive problem-solving, and continuous improvement. If you want to grow your technical leadership skills while making a tangible difference in how systems perform at scale, we would love to have you join us.
BASIC QUALIFICATIONS
- Bachelor's degree
- Experience in automating, deploying, and supporting large-scale infrastructure
- Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust
- Experience with Linux/Unix
- Experience with CI/CD pipelines build processes
- Current, active US Government Security Clearance of TS/SCI with Polygraph
PREFERRED QUALIFICATIONS
- Experience with distributed systems at scale
The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.
USA, MD, Jessup - 129,200.00 - 174,800.00 USD annually