Production System Engineer Project Intern (Server DevOps) - 2027 Start

    Highlights

    We work at the center of a large network of teams, including hardware design, data center operations, field maintenance, supply chain, and automation engineering, to keep our global infrastructure running reliably and efficiently at scale. Collaborate with engineers across hardware, ops, platform, and supply chain teams on real projectsMinimum Qualifications: Currently pursuing a Bachelor's or Master's student in CS, Computer Engineering, EE, IT, or a related field.

    Numbers & Facts

    LocationSan Jose, CA

    Description

    Our Server Fleet Operations team manages the full lifecycle of servers powering our self-built data centers across the United States and Europe: from new hardware rollout to decommissioning. We work at the center of a large network of teams, including hardware design, data center operations, field maintenance, supply chain, and automation engineering, to keep our global infrastructure running reliably and efficiently at scale.

    As a Project Intern, you will contribute to impactful short-term projects and gain hands-on experience in a fast-paced, professional environment. This internship offers the opportunity to develop practical skills, apply your knowledge to real-world challenges, and explore your career interests.

    Applications are reviewed on a rolling basis, so we encourage you to apply early.

    What You'll Do

    • Get hands-on with server deployment, monitoring, and maintenance across CPU and GPU fleets
    • Write scripts and small tools (Python, Bash, or similar) to automate repetitive operational tasks
    • Learn to troubleshoot real Linux issuesL OS, hardware, storage, networking, and performance
    • Explore GPU and AI infrastructure, and contribute to the tools that keep it reliable
    • Dig into infrastructure data and metrics to spot trends and improvement opportunities
    • Experiment with applying AI/LLMs to infrastructure troubleshooting and operations
    • Support incident investigations and root-cause analysis alongside senior engineers
    • Help write and improve documentation, runbooks, and internal knowledge bases
    • Collaborate with engineers across hardware, ops, platform, and supply chain teams on real projectsMinimum Qualifications:
    • Currently pursuing a Bachelor's or Master's student in CS, Computer Engineering, EE, IT, or a related field
    • Comfortable with basic Linux administration and troubleshooting
    • Some scripting/programming experience (Python, Bash, Go, or similar)
    • Basic understanding of hardware, OS concepts, networking, or distributed systems
    • Some exposure to server hardware, firmware (BIOS/UEFI), or system architecture

    Preferred Qualifications:

    • Hands-on experience from projects, labs, internships, open-source work, or competitions, academic or personal
    • Curiosity about AI agents, LLMs, or AI-powered automation
    • Strong problem-solving instincts and comfort diving into unfamiliar technical territory
    • Good communication skills and enjoys working as part of a team

    Similar Jobs

    See more jobs