Role at a glance
- Job function
-
Software Engineering & IT Internet of Things (IoT) Engineering DevOps & Site Reliability Engineering Systems Software Engineering
- Salary
- $143.7K – $194.4K/yr
- Location
- Seattle, Washington, United States
- Work arrangement
- On-site
- Employment
- Internship
- Education
- Bachelor's
Spotted an issue?
We’ll check it against the original posting.
Qualifications
Required
- 3+ years of non-internship professional software development experience
- Bachelor's degree or equivalent
- Knowledge of professional software engineering & best practices for full software development life cycle, including coding standards, software architectures, code reviews, source control management, continuous deployments, testing,...
- • Strong Linux and systems-level skills including production debugging across the hardware–software boundary
- • Practical distributed-systems experience including state machines, idempotency, reconciliation, retries, and persistence
- • Ability to design offline-first and reconnect-and-reconcile behavior for edge systems operating without guaranteed connectivity
- • Experience developing cloud-side services that orchestrate, monitor, and coordinate with fleets of edge devices
- • Track record of independently owning a substantial subsystem end to end from design through production
Preferred
- • Experience with containers at the edge (containerd, K3s, KubeEdge) and judgment on where full Kubernetes is inappropriate for constrained environments
- • Hands-on experience with heterogeneous compute platforms (x86, ARM, GPU, NPU) and accelerator or model versioning
- • Exposure to robotics frameworks such as ROS 2, DDS middleware, NVIDIA Jetson, or Isaac
- • Experience with secure OTA frameworks (TUF, Uptane) and hardware security primitives (TPM, HSM, secure boot)
- • Background in reliability-critical domains such as automotive, autonomous vehicles, industrial IoT, avionics, or telecom edge
- • Experience with large artifact or model distribution including caching strategies and bandwidth optimization
- • Familiarity with AWS IoT, AWS IoT Greengrass, or cloud-based device management services
About the role
Original posting provided by amazon
AWS IoT is building the infrastructure to deliver AI from the cloud to the physical world. Our Physical AI team is developing services that provision, deploy, monitor, and secure AI software across every processor inside autonomous machines including robots, vehicles, drones, agricultural equipment, and industrial workcells operating in environments with limited or no cloud connectivity.
We are looking for a Software Development Engineer II to own and drive the design and delivery of core subsystems in our edge-first fleet management platform for Physical AI. You will design and build systems that operate reliably without cloud connectivity handling offline provisioning, state reconciliation on reconnect, staged fleet-wide rollouts, atomic rollback across multi-processor machines, and secure over-the-air (OTA) delivery of AI models and software to heterogeneous processor topologies. Your systems must be correct when disconnected, consistent when reconnected, and resilient when partially connected. You will also develop cloud-side services that orchestrate, monitor, and coordinate with fleets of edge devices, ensuring seamless interaction between cloud control planes and disconnected or intermittently connected hardware. You will write high-performance code, working across the hardware–software boundary to debug issues that span operating systems, runtimes, networks, and physical devices. You will independently own substantial subsystems end to end from writing the design document through implementation, testing, and production operation.
You will also mentor junior engineers, raise the bar on design quality, and help shape the technical direction of the platform. This is an opportunity to build foundational infrastructure for a new category of AWS services delivering intelligence to the physical world at scale.
Key job responsibilities
• Design, build, and operate core subsystems of an edge-first fleet management platform, spanning offline provisioning, state reconciliation on reconnect, staged fleet-wide rollouts, and atomic rollback across multi-processor machines
• Develop cloud-side services that orchestrate, monitor, and coordinate with fleets of edge devices, ensuring reliable interaction between cloud control planes and disconnected or intermittently connected hardware
• Build secure over-the-air (OTA) delivery pipelines for AI models and software across heterogeneous processor topologies
• Write and review high-performance, production-quality code, and debug issues that span operating systems, runtimes, networks, and physical devices
• Author design documents for new subsystems and drive them independently from proposal through implementation, testing, and production operation
• Define and uphold correctness guarantees for disconnected, reconnecting, and partially connected system states
• Participate in on-call rotation, troubleshoot production issues, and drive root-cause fixes for the platform
• Mentor junior engineers and raise the engineering bar through code review, design review, and technical guidance
• Collaborate with product, hardware, and adjacent platform teams to define requirements and integration points for new fleet management capabilities
• Contribute to the technical roadmap and architectural direction of the platform as it scales to new device types and processor topologies
A day in the life
Your day starts by checking dashboards — on-call weeks bring pages like an OTA rollback spike; sprint weeks lead into standup with a feature update. Mid-morning often means design review, debating edge cases like reconciling conflicting state after a multi-day disconnect. Sprint weeks give focused coding blocks — building rollback logic across processors, pairing on race conditions. On-call weeks fill that time with live debugging across hardware and software, plus maintenance like patching dependencies. Afternoons bring CR reviews and mentoring a junior engineer through root-causing a bug. You close by updating Taskei and reviewing a peer's design doc.
About the team
ABOUT AWS:
Diverse Experiences
Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying.
Why AWS
Amazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.
Work/Life Balance
We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why flexible work hours and arrangements are part of our culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.
Inclusive Team Culture
Here at AWS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity and AmazeCon conferences, inspire us to never stop embracing our uniqueness.
Mentorship and Career Growth
We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.
Basic Qualifications
- 3+ years of non-internship professional software development experience- Bachelor's degree or equivalent
- Knowledge of professional software engineering & best practices for full software development life cycle, including coding standards, software architectures, code reviews, source control management, continuous deployments, testing, and operational excellence
- • Strong Linux and systems-level skills including production debugging across the hardware–software boundary
- • Practical distributed-systems experience including state machines, idempotency, reconciliation, retries, and persistence
- • Ability to design offline-first and reconnect-and-reconcile behavior for edge systems operating without guaranteed connectivity
- • Experience developing cloud-side services that orchestrate, monitor, and coordinate with fleets of edge devices
- • Track record of independently owning a substantial subsystem end to end from design through production
- • Ability to author solid design documents and make sound architectural decisions without being handed the architecture
- • Experience with fleet management or OTA update systems including staged rollout, rollback, pause-resume, and dependency handling
Preferred Qualifications
- • Experience with containers at the edge (containerd, K3s, KubeEdge) and judgment on where full Kubernetes is inappropriate for constrained environments- • Hands-on experience with heterogeneous compute platforms (x86, ARM, GPU, NPU) and accelerator or model versioning
- • Exposure to robotics frameworks such as ROS 2, DDS middleware, NVIDIA Jetson, or Isaac
- • Experience with secure OTA frameworks (TUF, Uptane) and hardware security primitives (TPM, HSM, secure boot)
- • Background in reliability-critical domains such as automotive, autonomous vehicles, industrial IoT, avionics, or telecom edge
- • Experience with large artifact or model distribution including caching strategies and bandwidth optimization
- • Familiarity with AWS IoT, AWS IoT Greengrass, or cloud-based device management services
Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.
Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.
The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.
USA, WA, Seattle - 143,700.00 - 194,400.00 USD annually
About the company
amazon
Large Enterprise
Amazon is a global leader in e-commerce and cloud computing, founded in 1994 by Jeff Bezos. Initially starting as an online bookstore, it has since expanded its offerings to include a vast range of products and services, including electronics, fashion, and digital content. With Amazon Web Services (AWS), the company also provides powerful cloud solutions to businesses around the world. Known for its innovation, customer-centric approach, and commitment to operational efficiency, Amazon continues to shape the future of retail and technology, consistently seeking new ways to enhance customer experiences.