amazon

amazon

Posted via Amazon Jobs

Systems Development Engineer II, Managed Operations (MO); AWS Operations Management (AWSOM) Team

Posted Sep 28, 2026

Role at a glance

Job function
Software Engineering & IT DevOps & Site Reliability Engineering
Salary
Not Disclosed
Location
Herndon, Virginia, United States
Work arrangement
On-site
Experience
3+ years of administrative experience in networking, storage systems, operating systems and hands-on systems engineering experience

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

Systems Development Engineers balance day-to-day operations of AWS software systems with long-term engineering improvements intended to reduce operational toil. The role works on the availability, reliability, performance, and efficiency of AWS Regions.

What You'll Do

  • Investigate production issues, perform root-cause analysis, and implement fixes to restore system health.
  • Identify patterns across incidents and collaborate on solutions that address broader classes of problems.
  • Evaluate Service Level Objectives, work with stakeholder teams on thresholds, and update infrastructure as code to keep monitoring...
  • Build automation and fleet-management systems, including workload migrations to more optimal hardware.
  • Develop monitoring, diagnostics, repair, and self-healing features and improve system-management tools and processes.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Basic qualifications: 3+ years of administrative experience in networking, storage systems, operating systems and hands-on systems engineering; programming experience with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, or Rust; Linux/Unix experience; and experience with CI/CD pipelines, DevOps practices, and Generative AI technologies. The posting also describes GenAI-assisted development, testing, documentation research, and responsible-use practices.

Required

  • 5+ years of administrative experience in networking, storage systems, operating systems and hands-on systems engineering experience
  • 2+ years of non-internship professional software development experience
  • Experience with automation or working with scripting languages like Python, Java, Perl, PHP, Ruby, Bash, Shell, or equivalent
  • Knowledge of networking protocols, security principles, and cloud technologies (particularly AWS)
  • Strong problem-solving and analytical capabilities

Original job description

Content provided by the employer

Do you love decomposing problems to develop products that impact millions of people around the world? Would you enjoy identifying, defining, and building software solutions that revolutionize how businesses operate?

Would you enjoy diving deep into operating and improving some of the largest software systems humanity has ever built? Do the challenges that come of driving technical, business, and cultural change to improve the reliability, performance, and efficiency excite you?

The AWS Managed Operations (MO) organization was founded iwith the objective to reduce operational load and toil through long-term engineering & Generative Artificial Intelligence (GenAI) projects. Managed Operations (MO) is building the best-in-class engineering and operations team that will own the day-to-day operations for AWS Regions; improving the availability, reliability, latency, performance and efficiency to operate AWS regions.

Amazon is looking for highly motivated Systems Development Engineers who can balance the day-to-day operations of AWS’ software systems with long-term software engineering to reduce operational toil. We need engineers who enjoy constantly learning and diving deep into the wide range of systems and technologies that make up one of the world’s largest cloud providers.

Key job responsibilities
Our engineers collaborate across diverse teams, projects, and environments to have a firsthand impact on our global customer base. You’ll bring a passion for innovation, data, search, analytics, and distributed systems.

You’ll also:

- Build solutions to innovate our solutions to agentic-first interfaces
- Solve challenging technical problems, often ones not solved before, at every layer of the stack.
 - Define system requirements, participate in the development and delivery of operability-related features such as system health monitoring, diagnostics, repair, and other self-healing automation
- Develop or further existing application and system management tools and processes that reduce manual efforts and increase overall efficiency

Generative AI & Accelerated Development

- Demonstrated ability to leverage Generative AI tools and AI-assisted development environments to accelerate prototyping, development, validation, and testing workflows
- Experience using AI-powered code generation and review tools to rapidly iterate on infrastructure automation scripts, configuration templates, and systems tooling and to automate investigative, diagnostic, or operational tasks
- Ability to apply GenAI capabilities to accelerate test case generation, integration testing, and validation of complex infrastructure components — reducing cycle times without sacrificing quality or security rigor
- Comfort using GenAI tools to rapidly synthesize technical documentation, compliance frameworks, and architectural patterns — accelerating research and decision-making in ambiguous problem spaces
- Awareness of the responsible use of GenAI in security-sensitive environments, including understanding of data handling boundaries, model limitations, and appropriate human-in-the-loop validation practices



A day in the life
You will split your time approximately 60/40 between operating production systems and driving long-term improvements to the reliability, availability, and performance of those systems.

What a Typical Week Looks Like:

- Root cause analysis & remediation — Investigate production issues such as deployment failures, identify underlying bugs, and implement fixes to restore system health.
- Systems-level problem solving — Identify patterns across individual / service incidents, design solutions that address entire classes of problems, and collaborate with your team to refine those designs.
- SLO stewardship — Evaluate the effectiveness of Service Level Objectives, partner with stakeholder teams to validate thresholds, and update infrastructure as code (IaC) to keep monitoring meaningful and actionable.
- Systems automation for operational excellence — Design and build systems that improve fleet management—such as safely migrating workloads to more optimal hardware types—delivering measurable gains in performance.

About the team
About AWS

Amazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.

Why AWS?

Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we’re building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future.

AWS Infrastructure Services (AIS)

AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain — and we’re looking for talented people who want to help.

Diverse Experiences

AWS values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying.

Inclusive Team Culture

AWS values curiosity and connection. Our employee-led and company-sponsored affinity groups promote inclusion and empower our people to take pride in what makes us unique. Our inclusion events foster stronger, more collaborative teams. Our continual innovation is fueled by the bold ideas, fresh perspectives, and passionate voices our teams bring to everything we do.

Mentorship & Career Growth

We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.

Work/Life Balance

We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.

Basic Qualifications

- 3+ years of administrative experience in networking, storage systems, operating systems and hands-on systems engineering experience
- Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust
- Experience with Linux/Unix
- Experience with CI/CD pipelines, DevOps practices, and Generative AI technologies

Preferred Qualifications

- 5+ years of administrative experience in networking, storage systems, operating systems and hands-on systems engineering experience
- 2+ years of non-internship professional software development experience
- Experience with automation or working with scripting languages like Python, Java, Perl, PHP, Ruby, Bash, Shell, or equivalent
- Knowledge of networking protocols, security principles, and cloud technologies (particularly AWS)
- Strong problem-solving and analytical capabilities

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.



USA, VA, Herndon - 129,200.00 - 174,800.00 USD annually
amazon

About the company

amazon

Large Enterprise

Amazon is a global leader in e-commerce and cloud computing, founded in 1994 by Jeff Bezos. Initially starting as an online bookstore, it has since expanded its offerings to include a vast range of products and services, including electronics, fashion, and digital content. With Amazon Web Services (AWS), the company also provides powerful cloud solutions to businesses around the world. Known for its innovation, customer-centric approach, and commitment to operational efficiency, Amazon continues to shape the future of retail and technology, consistently seeking new ways to enhance customer experiences.