Senior Cloud Hardware Engineer, AI/ML Storage Servers

Amazon

Denver, CO

JOB DETAILS
SKILLS
Amazon Web Services (AWS), Artificial Intelligence (AI), CPU (Central Processing Unit), Cloud Computing, Communication Skills, Computer Engineering, Computer Firmware, Computer Servers, Customer Relations, Design Verification, Disability Accommodations, Distributed Computing, Electrical Engineering, English Language, Establish Priorities, Failure Analysis, Functional Testing, GPU (Graphics Processing Unit), Hardware Debugging, Hardware Development, Identify Issues, Industrial Engineering, Large-Scale Systems, Leading Edge Technology, Manufacturing, Manufacturing/Industrial Processes, Mechanical Design, Mechanical Testing, Memory Hardware, Mentoring, Needs Assessment, Network Architecture/Engineering, Network Operations Center, Onboarding, Operations Management, Original Design Manufacturer (ODM), PCI Express (PCI-E), Presentation/Verbal Skills, Problem Solving Skills, Process Improvement, Product Development, Product Engineering, Product Management, Production Control, Production Systems, Quality Monitoring, Requirements Management, Research & Development (R&D), Safety Standards, Server Hardware, Server Programming/Applications, Signal Integrity, Software Design, Solid State Drive (SSD), Startup, Supply Chain, Supply Chain Management, System Architecture, System Integration (SI), Systems Reliability, Team Lead/Manager, Team Player, Technical Leadership, Test Plan/Schedule, Testability, Testing, Time Management, Topology, Vehicle Fleets, Verification Plans, Web Services, Writing Skills, x86 Processors
LOCATION
Denver, CO
POSTED
Today

DescriptionApplication deadline: May 22, 2026AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we're the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain — and we're looking for talented people who want to help.You'll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You'll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you'll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.Amazon Web Services (AWS) Hardware Engineering team creates compute and storage server designs for Amazon's web services. Our engineers work with leading edge technologies, solve challenging problems, influence the industry's roadmaps, and develop new and unique solutions that are ahead of the pack. We work in an environment that fosters innovation and creativity. We encourage and invest in new directions and new ideas that will serve our customers better. We sometimes fail, but we learn from such failures and we improve. It is because of the team and our constant focus on the customers, that we are able to develop creative and new designs that set the standards on performance, quality, cost, and operational excellence.What you will doAs a member of the Server Hardware Engineering team you will own and lead the design and development of server products utilizing a wide range of design and manufacturing partners.You will work closely with our customers to understand their technical needs and business goals, leveraging your experience with server design and the knowledge of various teams to architect the solutions that we will deploy at scale.To deliver your products you will work with an interdisciplinary team of component, firmware, test, qualification, and integration engineers, and lead our design and manufacturing partners to bring these servers to the data center. After launch you will oversee the fleet of servers you develop, monitoring their quality and how they are meeting the customer requirements.This is a fast-paced, intellectually challenging position, and you'll work with thought leaders in multiple technology areas. You'll have high standards for yourself and everyone you work with, and you'll be constantly looking for ways to improve your products' performance, quality, and cost. We're changing an industry, and we want individuals who are ready for this challenge and want to reach beyond what is possible today.Key job responsibilitiesLead technical solutions for complex storage and/or accelerator server and rack system architectural challengesOwn end-to-end system reliability, proactively identifying and resolving deficiencies before customer impactDesign and implement solutions to address system-level issues at large scaleDecompose complex server system problems (testability, reliability, diagnostics) into deliverable tasks and featuresApply expertise across hardware, software, system design, x86 architecture, processes, and operationsCollaborate with hardware, software, manufacturing, supply chain and product management teamsDevelop and implement diagnostic tools and monitoring solutions for production systemsDebug complex system failures in time sensitive settingsAbout the teamAmazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.Amazon Web Services (AWS) values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying.We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud.Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences.We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.Basic QualificationsExperience in developing functional specifications, design verification plans and functional test proceduresExperience in server technologies such as, thermal, mechanical, power, and signal integrityBachelor's degree or above in electrical engineering, computer engineering, or equivalent5+ years of Design/Innovation, research & development, manufacturing, process, industrial engineering, or related experienceExperience leading process improvement, systems development, and project managementExperience in English-language communication skills, both written and verbal7+ years of equivalent experienceIn depth expertise in one or more server technologies such as Thermal / Mechanical design, high speed bus design and signal integrity, failure analysis, server components (e.g. CPU, GPU, SSDs, memory), BIOS, BMC, and networkingPreferred QualificationsExperience working with technical and product stakeholders to define requirements, prioritize features, and influence product roadmapsExperience leading engineering teams as a mentor or tech lead, or experience with general troubleshooting/debugging of hardware3+ years of new hardware product development experience, e.g. server, storage, networking, or large-scale distributed systems experienceIn-depth expertise in one or more server technologies: thermal/mechanical design, high-speed bus design and signal integrity, failure analysis, server components (CPU, GPU, SSDs, memory), BIOS, BMC, and networkingExperience developing and executing test procedures for mechanical or electrical systems/componentsExperience working with ODMs/manufacturer through the product development and manufacturing lifecycleExperience building predictive failure detection or proactive remediation systems at fleet scaleExperience with storage/compute/GPU/accelerator platforms including integration, diagnostics, or performance validationFamiliarity with PCIe topology, NVLink, NVMe, and accelerator interconnectsExperience with large-scale datacenter or cloud environmentsAmazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner.The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at .USA, CA, Cupertino - 183,000.00 - 247,600.00 USD annuallyUSA, WA, Seattle - 159,200.00 - 215,300.00 USD annually#J-18808-Ljbffr

About the Company

A

Amazon

At Amazon, we don’t wait for the next big idea to present itself. We envision the shape of impossible things and then we boldly make them reality. So far, this mindset has helped us achieve some incredible things. Let’s build new systems, challenge the status quo, and design the world we want to live in. We believe the work you do here will be the best work of your life.

Wherever you are in your career exploration, Amazon likely has an opportunity for you. Our research scientists and engineers shape the future of natural language understanding with Alexa. Fulfillment center associates around the globe send customer orders from our warehouses to doorsteps. Product managers set feature requirements, strategy, and marketing messages for brand new customer experiences. And as we grow, we’ll add jobs that haven’t been invented yet.

It’s Always Day 1
At Amazon, it’s always “Day 1.” Now, what does this mean and why does it matter? It means that our approach remains the same as it was on Amazon’s very first day – to make smart, fast decisions, stay nimble, invent, and stay focused on delighting our customers. In our 2016 shareholder letter, Amazon CEO Jeff Bezos shared his thoughts on how to keep up a Day 1 company mindset. “Staying in Day 1 requires you to experiment patiently, accept failures, plant seeds, protect saplings, and double down when you see customer delight,” he wrote. “A customer-obsessed culture best creates the conditions where all of that can happen.” You can read the full letter here

Our Leadership Principles
Our Leadership Principles help us keep a Day 1 mentality. They aren’t just a pretty inspirational wall hanging. Amazonians use them, every day, whether they’re discussing ideas for new projects, deciding on the best solution for a customer’s problem, or interviewing candidates. To read through our Leadership Principles from Customer Obsession to Bias for Action, visit https://www.amazon.jobs/principles
COMPANY SIZE
10,000 employees or more
INDUSTRY
Retail
FOUNDED
1994
WEBSITE
http://Amazon.com/militaryroles