Skip to main content
A

SDE II , AWS IoT Fleet Management

Amazon Web Services (AWS)
3 hours ago
Full-time
On-site
Seattle, WA, United States
LinkedIn

Description AWS Applied AI Solutions (AAIS) is building toward a future where every business innovates with Amazon AI teammates. To get there, we build AI solutions that improve human capabilities and transform entire business functions. We create end-to-end products that surprise and delight out-of-the-box, making complex things easy and hard things possible, with no cloud experience required. We start with customers who embrace the future and build bridges to meet the rest where they are. We pursue ambitious opportunities with conviction, and we are looking for builders who share that mindset. AWS IoT is building the infrastructure to deliver AI from the cloud to the physical world. Our Physical AI team is developing services that provision, deploy, monitor, and secure AI software across every processor inside autonomous machines including robots, vehicles, drones, agricultural equipment, and industrial workcells operating in environments with limited or no cloud connectivity. We are looking for a Software Development Engineer II to own and drive the design and delivery of core subsystems in our edge-first fleet management platform for Physical AI. You will design and build systems that operate reliably without cloud connectivity handling offline provisioning state reconciliation on reconnect staged fleet-wide rollouts atomic rollback across multi-processor machines and secure over-the-air (OTA) delivery of AI models and software to heterogeneous processor topologies. Your systems must be correct when disconnected consistent when reconnected and resilient when partially connected. You will also develop cloud-side services that orchestrate, monitor, and coordinate with fleets of edge devices ensuring seamless interaction between cloud control planes and disconnected or intermittently connected hardware. You will write high-performance code working across the hardware-software boundary to debug issues that span operating systems, runtimes, networks, and physical devices. You will independently own substantial subsystems end to end from writing the design document through implementation testing and production operation. You will also mentor junior engineers raise the bar on design quality and help shape the technical direction of the platform. This is an opportunity to build foundational infrastructure for a new category of AWS services delivering intelligence to the physical world at scale. Key job responsibilities Design, build, and operate core subsystems of an edge-first fleet management platform spanning offline provisioning state reconciliation on reconnect staged fleet-wide rollouts and atomic rollback across multi-processor machines Develop cloud-side services that orchestrate, monitor, and coordinate with fleets of edge devices ensuring reliable interaction between cloud control planes and disconnected or intermittently connected hardware Build secure over-the-air (OTA) delivery pipelines for AI models and software across heterogeneous processor topologies Write and review high-performance production-quality code and debug issues that span operating systems, runtimes, networks, and physical devices Author design documents for new subsystems and drive them independently from proposal through implementation testing and production operation Define and uphold correctness guarantees for disconnected reconnecting and partially connected system states Participate in on-call rotation troubleshoot production issues and drive root-cause fixes for the platform Mentor junior engineers and raise the engineering bar through code review, design review, and technical guidance Collaborate with product hardware and adjacent platform teams to define requirements and integration points for new fleet management capabilities Contribute to the technical roadmap and architectural direction of the platform as it scales to new device types and processor topologies A day in the life Your day starts by checking dashboards - on-call weeks bring pa…

View the full posting and apply