Amazon.com Inc logo

Technical Program Manager, Availability Serviceability Team

Amazon.com Inc

  • Herndon, VA
  • 6 days ago

    Highlights

    The AWS Data Center Availability Serviceability team works with hundreds of AWS data centers globally to deliver the highest quality and lowest cost availability, capacity, and scaling results for our customers. Why you will love it: You will own a program that is genuinely novel -- building something that didn"t exist two years ago, using AI tooling that is still maturing, against a problem (fleet-wide MBOM completeness) that no one has solved at this scale.

    Numbers & Facts

    LocationHerndon, VA
    IndustryRetail
    Company Size10,000 employees or more
    Year Founded1994
    Websitehttp://Amazon.com/militaryroles

    Description

    The AWS Data Center Availability Serviceability team works with hundreds of AWS data centers globally to deliver the highest quality and lowest cost availability, capacity, and scaling results for our customers. We standardize operations globally by delivering tools, policy, processes, and procedures to our internal teams.

    Serviceability is the ability to maintain, repair, and manage infrastructure assets efficiently throughout their lifecycle -- encompassing maintenance standards, spares management, training and certifications, operational readiness, and the systematic use of data and automation to prevent disruptions and drive continuous improvement.

    We are seeking a Technical Program Manager to own the MBOM program as part of the AWS Data Center Availability Serviceability Critical Spares team. MBOMs define which parts make up each piece of critical infrastructure equipment and are foundational to programmatic spares planning and fulfillment -- enabling proactive stocking decisions, criticality assessments, and repair readiness across the global fleet.

    What you will do:

    You will own the development and delivery of a multi-year roadmap to establish comprehensive MBOM coverage across the global AWS fleet. Working with stakeholders across Field Engineering, DCEO, vendors, and lifecycle data teams, you will scale AI-powered pipelines that extract MBOM data from vendor documentation, define per-equipment MBOM standards, and design the processes for establishing MBOMs where source documentation does not exist. Your work sets the foundation for programmatic spares identification, criticality assessment, and fulfillment -- ensuring the right parts are available at the right locations before failures occur. You will build the mechanisms that keep MBOM data current as the fleet grows, diversifies into new cooling technologies, and introduces new vendors.

    Why it matters:

    Spares readiness starts with knowing what"s inside the equipment. MBOM is the data layer that connects equipment design to spares strategy to repair outcomes. With comprehensive MBOM coverage, we can programmatically identify critical parts, make proactive stocking decisions, and reduce time to repair -- directly improving data center availability for customers.

    Why you will love it:

    You will own a program that is genuinely novel -- building something that didn"t exist two years ago, using AI tooling that is still maturing, against a problem (fleet-wide MBOM completeness) that no one has solved at this scale. You will operate in high ambiguity with real ownership to define what good looks like. You will work across PLM systems, enterprise asset management, spares strategy, and physical equipment -- giving you breadth that most TPM roles don"t offer. And your output has direct, measurable impact: every MBOM you complete makes the fleet more repairable.

    Key job responsibilities

    • Own MBOM program strategy, execution, and completeness metrics across the global fleet
    • Operate and scale AI-assisted pipelines that convert vendor documentation into governed MBOM records
    • Define and execute the process for establishing MBOMs where vendor documentation is limited or unavailable
    • Ensure MBOM data flows into criticality assessments and stocking decisions within the Critical Spares program
    • Define per-equipment MBOM standards in partnership with Field Engineering and equipment owners
    • Build mechanisms to maintain MBOM accuracy as new equipment deploys and the fleet evolves (including liquid cooling)
    • Drive cross-functional coordination across Availability, Field Engineering, DCEO, and lifecycle data teams
    • Use data and metrics to measure program health, prioritize effort, and communicate progress
    • Write narratives (one to six pages) and present to Director-level leadership
    • Travel up to 10%

    A day in the life

    You will split your time between program execution and cross-functional coordination. On any given day you might be reviewing AI pipeline output quality with the tooling team, defining MBOM standards for a new equipment type with Field Engineering, working with a vendor to obtain documentation for legacy equipment, analyzing coverage gaps to prioritize the next tranche of MBOMs, or presenting program status and trade-off decisions to leadership. You will operate at both the strategic level - setting the roadmap for fleet-wide MBOM completeness - and the tactical level - digging into specific equipment families where data doesn"t exist.

    About Company

    At Amazon, we don’t wait for the next big idea to present itself. We envision the shape of impossible things and then we boldly make them reality. So far, this mindset has helped us achieve some incredible things. Let’s build new systems, challenge the status quo, and design the world we want to live in. We believe the work you do here will be the best work of your life.

    Wherever you are in your career exploration, Amazon likely has an opportunity for you. Our research scientists and engineers shape the future of natural language understanding with Alexa. Fulfillment center associates around the globe send customer orders from our warehouses to doorsteps. Product managers set feature requirements, strategy, and marketing messages for brand new customer experiences. And as we grow, we’ll add jobs that haven’t been invented yet.

    It’s Always Day 1
    At Amazon, it’s always “Day 1.” Now, what does this mean and why does it matter? It means that our approach remains the same as it was on Amazon’s very first day – to make smart, fast decisions, stay nimble, invent, and stay focused on delighting our customers. In our 2016 shareholder letter, Amazon CEO Jeff Bezos shared his thoughts on how to keep up a Day 1 company mindset. “Staying in Day 1 requires you to experiment patiently, accept failures, plant seeds, protect saplings, and double down when you see customer delight,” he wrote. “A customer-obsessed culture best creates the conditions where all of that can happen.” You can read the full letter here

    Our Leadership Principles
    Our Leadership Principles help us keep a Day 1 mentality. They aren’t just a pretty inspirational wall hanging. Amazonians use them, every day, whether they’re discussing ideas for new projects, deciding on the best solution for a customer’s problem, or interviewing candidates. To read through our Leadership Principles from Customer Obsession to Bias for Action, visit https://www.amazon.jobs/principles

    Similar Jobs

    See more jobs