Amazon.com Inc logo

Senior Technical Program Manager (SW/FW) , AWS Hardware Engineering

Amazon.com Inc

  • Seattle, WA
  • 4 days ago

    Highlights

    Writing down a gate that only existed in people"s heads, replacing a manual status roll-up with something automated, tightening a handoff between two teams that keeps slipping, updating the release schedule as reality changes. The AWS Hardware Engineering team designs the servers behind AWS, from general purpose compute and storage to the GPU and AI accelerator systems used for machine learning and generative AI.

    Numbers & Facts

    LocationSeattle, WA
    IndustryRetail
    Company Size10,000 employees or more
    Year Founded1994
    Websitehttp://Amazon.com/militaryroles

    Description

    Every server in an AWS data center runs firmware, and someone has to decide how a new version of it safely reaches millions of machines. That is this role: you will own how firmware ships across the AWS server fleet, and you will design the mechanisms that make it predictable and safe at that scale.

    We build the software that lives closest to the hardware, including BMC, BIOS, and the controllers that manage power, cooling, sensors, and health on every server and rack we deploy. Writing that firmware is only half the problem. Getting a new version onto millions of machines, in the right order, at the right time, without disrupting a single customer, is the other half. That half is yours.

    Your real product is the mechanisms. You will design the repeatable machinery that gets every release to the finish line without heroics: a release calendar the whole organization plans against, quality gates with clear criteria a build must clear before it goes anywhere near production, a staged rollout model that limits exposure if something goes wrong, and one shared view of where every release stands so nobody has to ask. You will define how teams hand off to each other, how a blocker gets escalated and by when, and how we know a rollout is actually healthy rather than just finished.

    Every server in an AWS data center runs firmware, and someone has to decide how a new version of it safely reaches millions of machines. That is this role: you will own how firmware ships across the AWS server fleet, and you will design the mechanisms that make it predictable and safe at that scale.

    We are looking for an experienced Senior Technical Program Manager to own firmware release and deployment across our firmware teams. You will work across firmware engineering, hardware engineering, manufacturing partners, service teams, and security to make firmware delivery predictable, repeatable, and safe at fleet scale.

    We build the software that lives closest to the hardware, including BMC, BIOS, and the controllers that manage power, cooling, sensors, and health on every server and rack we deploy. Writing that firmware is only half the problem. Getting a new version onto millions of machines, in the right order, at the right time, without disrupting a single customer, is the other half. That half is yours.

    Your real product is the mechanisms. You will design the repeatable machinery that gets every release to the finish line without heroics: a release calendar the whole organization plans against, quality gates with clear criteria a build must clear before it goes anywhere near production, a staged rollout model that limits exposure if something goes wrong, and one shared view of where every release stands so nobody has to ask. You will define how teams hand off to each other, how a blocker gets escalated and by when, and how we know a rollout is actually healthy rather than just finished.

    You will get to decide what good looks like here, then make it stick. That means finding the places where coordination happens by memory or by meeting and replacing them with something documented, measured, and ideally automated. It means using every release and every operational surprise as evidence, turning what you learn into a permanent change to the process or the tooling instead of a one-time fix. And it means holding a high bar with engineering teams who do not report to you, which you will earn through being right and being useful rather than through org chart authority.

    What makes this different from most program management jobs:

    • You will work where software meets physical hardware. A bad firmware rollout does not just throw an error, it can take a machine offline, so the bar for getting it right is unusually high.
    • Your scope is the whole fleet. Decisions you make about release process apply to millions of servers across every AWS region.
    • You are building the mechanisms, not inheriting them. There is real room to design how this works rather than administer someone else"s process.
    • You will get deep in the technology. This is a hands-on technical role with engineers, not a status-reporting role.

    If you like turning messy, high-stakes coordination into something that runs predictably, and you want to see your work land at a scale very few places can offer, we would like to talk to you. AWS engineers are shaping the way people use computers and designing the future of cloud computing technology, come help us make history!

    Key job responsibilities

    Own firmware release and deployment end to end, from a completed code change through qualification to a fully deployed fleet.

    Publish and maintain the release schedule that the firmware teams plan against, and drive agreement on scope, sequencing, and dependencies between them.

    Define the quality gates a release must pass at each stage, including hardware qualification and validation at our manufacturing partners, and hold teams to the entry and exit criteria.

    Build and run the reporting that shows where every release stands, how much of the fleet is on the current version, and what is blocking progress, and present that to senior leadership on a regular cadence.

    Own escalation for release blockers, including issues raised by manufacturing partners, and drive them to closure against a defined response time.

    Turn what we learn from each release and each operational issue into a lasting change to the process, the gates, or the tooling.

    Partner with firmware, hardware, manufacturing, security, and service teams to resolve competing priorities and keep releases moving.

    Find the coordination that happens manually today and replace it with automation.

    A day in the life

    You will spend time with the firmware engineering teams, going deep on what is in the next release, what changed, what worries them, and whether it is ready to move to the next stage. These are technical conversations, not status checks, and you will be expected to hold your own in them.

    You will spend time on releases already in flight. That means looking at how a staged rollout is progressing, deciding whether the data supports widening it, and pulling the handle to pause or roll back when it does not. When something fails qualification or a manufacturing partner reports a problem, you are the person who gets the right people on it and keeps it moving until it is closed.

    You will spend time building. Writing down a gate that only existed in people"s heads, replacing a manual status roll-up with something automated, tightening a handoff between two teams that keeps slipping, updating the release schedule as reality changes.

    And you will spend time communicating. Giving leadership a clear read on where releases stand and what the risks are, aligning teams whose priorities are in tension, and making the case for a decision that not everyone will initially agree with.

    About the team

    The AWS Hardware Engineering team designs the servers behind AWS, from general purpose compute and storage to the GPU and AI accelerator systems used for machine learning and generative AI. These systems run at high power and thermal density and are deployed in large clusters, where the health of a single server can affect a whole training job. Our firmware manages that hardware: power, cooling, sensors, health reporting, and its own updates. In this role you would own how that firmware reaches servers across the fleet.

    About Company

    At Amazon, we don’t wait for the next big idea to present itself. We envision the shape of impossible things and then we boldly make them reality. So far, this mindset has helped us achieve some incredible things. Let’s build new systems, challenge the status quo, and design the world we want to live in. We believe the work you do here will be the best work of your life.

    Wherever you are in your career exploration, Amazon likely has an opportunity for you. Our research scientists and engineers shape the future of natural language understanding with Alexa. Fulfillment center associates around the globe send customer orders from our warehouses to doorsteps. Product managers set feature requirements, strategy, and marketing messages for brand new customer experiences. And as we grow, we’ll add jobs that haven’t been invented yet.

    It’s Always Day 1
    At Amazon, it’s always “Day 1.” Now, what does this mean and why does it matter? It means that our approach remains the same as it was on Amazon’s very first day – to make smart, fast decisions, stay nimble, invent, and stay focused on delighting our customers. In our 2016 shareholder letter, Amazon CEO Jeff Bezos shared his thoughts on how to keep up a Day 1 company mindset. “Staying in Day 1 requires you to experiment patiently, accept failures, plant seeds, protect saplings, and double down when you see customer delight,” he wrote. “A customer-obsessed culture best creates the conditions where all of that can happen.” You can read the full letter here

    Our Leadership Principles
    Our Leadership Principles help us keep a Day 1 mentality. They aren’t just a pretty inspirational wall hanging. Amazonians use them, every day, whether they’re discussing ideas for new projects, deciding on the best solution for a customer’s problem, or interviewing candidates. To read through our Leadership Principles from Customer Obsession to Bias for Action, visit https://www.amazon.jobs/principles

    Similar Jobs

    See more jobs