Site Reliability Engineer

Obsidian Security

  • Palo Alto, CA
  • 30+ days ago

    Highlights

    We work closely with Engineering, Quality Engineering, and Customer Support to deliver end-to-end services that bring code to life and maintain our world-class SaaS security platform. The DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-performing production systems.

    Numbers & Facts

    LocationPalo Alto, CA

    Description

    About the DevOps / SRE Team

    The DevOps/SRE team at Obsidian ensures that engineering excellence translates into stable, scalable, and high-performing production systems. We work closely with Engineering, Quality Engineering, and Customer Support to deliver end-to-end services that bring code to life and maintain our world-class SaaS security platform.

    What You'll Do

    • Support and maintain the service quality of our customer-facing SaaS security platform
    • Address complex challenges around scalability, reliability, observability, and cost efficiency
    • Collaborate with Engineering teams to maintain and enhance Helm charts, application deployment, monitoring and CI/CD pipelines
    • Embed into the engineering team so that you understand the application deeply.
    • Define service verification strategies and implement them as part of the CI/CD process to meet SLAs
    • Improve developer experience by optimizing CI/CD workflows and performance
    • Participate in the on-call rotation, providing 24/7 support in coordination with our global SRE team
    • Monitor, debug, and optimize production infrastructure and services on AWS/GCP

    What We're Looking For

    • 3+ years of experience in a DevOps or SRE role supporting SaaS services on GCP and/or AWS
    • Bachelor's degree in Computer Science or related field
    • Strong proficiency in Kubernetes, microservices architecture, Helm, GitLab CI/CD, and ArgoCD, Prometheus, Grafana.
    • Programming experience in at least one language; Golang or Python preferred
    • Deep understanding of autoscaling, version upgrades, and cloud service optimization
    • Bonus if you're familiar with technologies like Kafka, Elasticsearch, PostgreSQL, ScyllaDB, Databricks, Dagster, Sentry, Kong

    Similar Jobs

    See more jobs