Amazon.com Inc logo

Software Development Engineer, Data Center Host Monitoring

Amazon.com Inc

  • Seattle, WA
  • 19 days ago

    Highlights

    We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. As an SDE on this team, you"ll own and evolve distributed services spanning availability monitoring, hardware telemetry pipelines, and anomaly detection and alarming - the systems that serve as the first line of detection when something goes wrong in a data center.

    Numbers & Facts

    LocationSeattle, WA
    IndustryRetail
    Company Size10,000 employees or more
    Year Founded1994
    Websitehttp://Amazon.com/militaryroles

    Description

    AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we're the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain - and we're looking for talented people who want to help.

    You'll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You'll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. And you'll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.

    Key job responsibilities

    • Design, build, and operate distributed services that are reliable, scalable, and maintainable at global scale
    • Own your services end-to-end - from architecture and implementation through deployment and production operations
    • Identify and drive improvements to system reliability, latency, and operational posture
    • Contribute to team engineering standards, code reviews, and technical direction
    • Mentor junior engineers and contribute to a culture of high engineering quality

    About the team

    Every server in every AWS data center is continuously generating signals - thermal readings, reachability status, power metrics - and it"s our team"s job to make sure that data is collected, processed, and acted on in real time. We build and operate the monitoring systems that give data center operations teams visibility into the health of AWS"s global infrastructure, and we power the alerting systems that notify operations staff when something needs their attention.

    As an SDE on this team, you"ll own and evolve distributed services spanning availability monitoring, hardware telemetry pipelines, and anomaly detection and alarming - the systems that serve as the first line of detection when something goes wrong in a data center. Your customers are internal: the engineers and operators who work hands-on with physical infrastructure every day. That means tight feedback loops, fast iteration, and the chance to sit down with the people actually using your systems to understand what"s working. Real Customer Obsession, not the abstract kind.

    About Company

    At Amazon, we don’t wait for the next big idea to present itself. We envision the shape of impossible things and then we boldly make them reality. So far, this mindset has helped us achieve some incredible things. Let’s build new systems, challenge the status quo, and design the world we want to live in. We believe the work you do here will be the best work of your life.

    Wherever you are in your career exploration, Amazon likely has an opportunity for you. Our research scientists and engineers shape the future of natural language understanding with Alexa. Fulfillment center associates around the globe send customer orders from our warehouses to doorsteps. Product managers set feature requirements, strategy, and marketing messages for brand new customer experiences. And as we grow, we’ll add jobs that haven’t been invented yet.

    It’s Always Day 1
    At Amazon, it’s always “Day 1.” Now, what does this mean and why does it matter? It means that our approach remains the same as it was on Amazon’s very first day – to make smart, fast decisions, stay nimble, invent, and stay focused on delighting our customers. In our 2016 shareholder letter, Amazon CEO Jeff Bezos shared his thoughts on how to keep up a Day 1 company mindset. “Staying in Day 1 requires you to experiment patiently, accept failures, plant seeds, protect saplings, and double down when you see customer delight,” he wrote. “A customer-obsessed culture best creates the conditions where all of that can happen.” You can read the full letter here

    Our Leadership Principles
    Our Leadership Principles help us keep a Day 1 mentality. They aren’t just a pretty inspirational wall hanging. Amazonians use them, every day, whether they’re discussing ideas for new projects, deciding on the best solution for a customer’s problem, or interviewing candidates. To read through our Leadership Principles from Customer Obsession to Bias for Action, visit https://www.amazon.jobs/principles

    Similar Jobs

    See more jobs