Site Reliability Engineer - Senior Staff

Ruckus

  • Sunnyvale, CA
  • 2 days ago

    Highlights

    How you'll help us connect the world:Ruckus Networks is looking for a customer-focused Senior Site Reliability Engineer (SRE) to help improve reliability, scalability, operational excellence, and customer experience across our cloud platform ecosystem. At Ruckus Networks, you will work on large-scale cloud networking platforms that support enterprise customers globally.

    Numbers & Facts

    LocationSunnyvale, CA

    Description

    In our “always on” world, we believe it's essential to have a genuine connection with the work you do. At Ruckus Networks, you will work on large-scale cloud networking platforms that support enterprise customers globally. You will help improve reliability, automation, observability, and customer experience while working with modern cloud and SRE technologies in a collaborative engineering environment.How you'll help us connect the world:Ruckus Networks is looking for a customer-focused Senior Site Reliability Engineer (SRE) to help improve reliability, scalability, operational excellence, and customer experience across our cloud platform ecosystem. This role is ideal for engineers who enjoy solving production problems, building automation, and improving platform reliability at scale. You will work on distributed systems powering cloud networking services used by customers globally in a fast‑paced environment. As part of the SRE organization, you will work closely with engineering, cloud operations, and support teams to improve platform stability, observability, automation, and operational readiness.Hybrid role: This is a hybrid role and needs to be on‑site at our Sunnyvale, CA office 3 days a week. No relocation or 3rd‑party agencies please.Key responsibilities:Reliability Engineering & OperationsOperate and improve highly available, scalable cloud services and infrastructureTroubleshoot production issues across applications, infrastructure, networking, databases, and cloud servicesImprove observability through metrics, logging, tracing, synthetic monitoring, and alertingHelp define and improve SLIs, SLOs, and operational health metricsParticipate in incident response and support Sev‑1/customer‑impacting eventsContribute to post‑incident reviews and long‑term reliability improvementsImprove operational processes, automation, and deployment safetyAutomation & EngineeringBuild operational tooling and automation using PythonImprove operational efficiency through automation and self‑service toolingSupport CI/CD improvements and deployment validation workflowsDevelop health checks, monitoring integrations, and operational diagnosticsCloud & InfrastructureSupport services running in Google Cloud Platform (GCP)Work with Kubernetes, containers, and cloud‑native platformsAnalyze scalability, performance, and resource utilizationCollaborate with software engineering teams on operational readiness and reliability improvementsObservability & MonitoringBuild dashboards, alerts, and telemetry pipelinesWork with observability platforms such as Prometheus, Grafana, OpenTelemetry, and ELKSupport monitoring and analytics platforms including ClickHouseImprove signal quality and reduce operational alert noiseDevelop synthetic monitoring focused on customer workflowsCollaborationPartner with Engineering, Product Management, Customer Support, and Cloud Operations teamsParticipate in architecture and operational readiness discussionsMentor junior engineers and contribute to SRE best practicesPromote operational excellence, ownership, and customer focusRequired qualifications:5+ years of experience in Site Reliability Engineering, DevOps, Cloud Infrastructure, or Production EngineeringStrong programming skills in PythonExperience with Linux systems administration and troubleshootingHands‑on experience with Google Cloud Platform (GCP)Experience with Kubernetes, containers, and cloud‑native infrastructureExperience troubleshooting distributed systems in production environmentsExperience with observability tools such as Prometheus, Grafana, Open Telemetry, or ELKFamiliarity with ClickHouse or large‑scale telemetry platformsUnderstanding of networking fundamentals, APIs, databases, and cloud architecturesExperience participating in production incident response and operational supportYou excite us if you have:Experience supporting SaaS or cloud platforms at scaleFamiliarity with Kafka or event‑driven architecturesExperience building automation and monitoring solutionsFamiliarity with wireless networking or enterprise networking platformsExperience improving operational processes and reliability practicesOur salary ranges consider a wide variety of factors, including but not limited to benchmarking by independent third‑party consultants, skill sets, years of experience, training, education, geography, and other business needs. Depending on experience, the range can be higher for candidates with exceptional experience and a demonstrated history of successful performance. This position's expected total compensation (base salary and commission range) is $118,000.00‑$170,000.00.The candidate will be rewarded with a comprehensive benefits package, including medical, dental, and vision plans, life and accidental death insurance, a 401(k) plan, and participation in the Company's Incentive Plan. Candidates starting with the Company will be eligible for eleven paid holidays in a full calendar year, two weeks of paid vacation (prorated based on start date), as well as other leave options.Why join us?Vistance Networks shapes the future of communications technology, pushing past what is possible. We deliver solutions that bring reliability and performance to a world always in motion. Our global team of innovators and employees are trusted advisors who listen to customers first, then deliver value.RUCKUS Networks delivers purpose‑driven enterprise networks that enable superior business outcomes in demanding environments. Our solutions combine AI‑powered automation, proactive network assurance, and context‑aware security, providing exceptional performance with simplified management.If you want to grow your career alongside bright, passionate, and caring people who strive to create what's next…come connect to your future at Vistance Networks.Vantage Networks is an Equal Opportunity Employer (EEO), including people with disabilities and veterans.#J-18808-Ljbffr

    Similar Jobs