| Location | Sunnyvale, CA |
| Salary | $118,000–$170,000 Per Year |
Overview Site Reliability Engineer - Senior StaffReq ID: 81736Location: Sunnyvale, California, United States, 94089In our ‘always on' world, we believe it's essential to have a genuine connection with the work you do.At Ruckus Networks, you will work on large-scale cloud networking platforms that support enterprise customers globally. You will help improve reliability, automation, observability, and customer experience while working with modern cloud and SRE technologies in a collaborative engineering environment.How You'll help us connect the world:Ruckus Networks is looking for a customer focused Senior Site Reliability Engineer (SRE) to help improve reliability, scalability, operational excellence, and customer experience across our cloud platform ecosystem.This role is ideal for engineers who enjoy solving production problems, building automation, and improving platform reliability at scale. You will work on distributed systems powering cloud networking services used by customers globally in fast paced environment.As part of the SRE organization, you will work closely with engineering, cloud operations, and support teams to improve platform stability, observability, automation, and operational readiness.THIS IS A HYBRID ROLE AND NEEDS TO BE ON-SITE AT OUR SUNNYVALE, CA OFFICE 3 DAYS A WEEK. NO RELOCATION OR 3RD PARTY AGENCIES PLEASEKey Responsibilities Operate and improve highly available, scalable cloud services and infrastructureTroubleshoot production issues across applications, infrastructure, networking, databases, and cloud servicesImprove observability through metrics, logging, tracing, synthetic monitoring, and alertingHelp define and improve SLIs, SLOs, and operational health metricsParticipate in incident response and support Sev-1/customer-impacting eventsContribute to post-incident reviews and long-term reliability improvementsImprove operational processes, automation, and deployment safetyAutomation & Engineering Build operational tooling and automation using PythonImprove operational efficiency through automation and self-service toolingSupport CI/CD improvements and deployment validation workflowsDevelop health checks, monitoring integrations, and operational diagnosticsCloud & Infrastructure Support services running in Google Cloud Platform (GCP)Work with Kubernetes, containers, and cloud-native platformsAnalyze scalability, performance, and resource utilizationCollaborate with software engineering teams on operational readiness and reliability improvementsObservability & Monitoring Build dashboards, alerts, and telemetry pipelinesWork with observability platforms such as Prometheus, Grafana, OpenTelemetry, and ELKSupport monitoring and analytics platforms including ClickHouseImprove signal quality and reduce operational alert noiseDevelop synthetic monitoring focused on customer workflowsCollaboration Partner with Engineering, Product Management, Customer Support, and Cloud Operations teamsParticipate in architecture and operational readiness discussionsMentor junior engineers and contribute to SRE best practicesPromote operational excellence, ownership, and customer focusRequired Qualifications 5+ years of experience in Site Reliability Engineering, DevOps, Cloud Infrastructure, or Production EngineeringStrong programming skills in PythonExperience with Linux systems administration and troubleshootingHands-on experience with Google Cloud Platform (GCP)Experience with Kubernetes, containers, and cloud-native infrastructureExperience troubleshooting distributed systems in production environmentsExperience with observability tools such as Prometheus, Grafana, Open Telemetry, or ELKFamiliarity with ClickHouse or large-scale telemetry platformsUnderstanding of networking fundamentals, APIs, databases, and cloud architecturesExperience participating in production incident response and operational supportYou Excite us if you have Experience supporting SaaS or cloud platforms at scaleFamiliarity with Kafka or event-driven architecturesExperience building automation and monitoring solutionsFamiliarity with wireless networking or enterprise networking platformsExperience improving operational processes and reliability practices#LI-RB1#LI-HYBRIDThe candidate will be rewarded with a comprehensive benefits package, including medical, dental, and vision plans, life and accidental death insurance, a 401(k) plan, and participation in the Company's Incentive Plan. Candidates starting with the Company will be eligible for eleven paid holidays in a full calendar year, two weeks of paid vacation (prorated based on start date), as well as other leave options.Compensation & EEO Our salary ranges consider a wide variety of factors, including but not limited to benchmarking by independent third-party consultants, skill sets, years of experience, training, education, geography, and other business needs. Depending on experience, the range can be higher for candidates with exceptional experience and a demonstrated history of successful performance. This position's expected total compensation (base salary and commission range) is $118,000.00-$170,000.00The candidate will be rewarded with a comprehensive benefits package, including medical, dental, and vision plans, life and accidental death insurance, a 401(k) plan, and participation in the Company's Incentive Plan. Candidates starting with the Company will be eligible for eleven paid holidays in a full calendar year, two weeks of paid vacation (prorated based on start date), as well as other leave options.Why Join Us? Vistance Networks shapes the future of communications technology, pushing past what is possible. We deliver solutions that bring reliability and performance to a world always in motion. Our global team of innovators and employees are trusted advisors who listen to customers first, then deliver value.Company RUCKUS Networks is an Equal Opportunity Employer (EEO), including people with disabilities and veterans.Learn more about how we're on a quest to connect the future and build what's next.Job Segment: Cloud, Senior Product Manager, Software Engineer, Linux, Network, Technology, Operations, Engineering#J-18808-Ljbffr