Site Reliability Engineer I

Consensus Cloud Solutions Inc

  • NY
  • 30+ days ago
  • Remote
  • $205,000–$212,500 Per Year

Highlights

Lead the design, development, and maintenance of secure, scalable, resilient, and cost-effective cloud infrastructure solutions on AWS through a DevOps approach, leveraging the existing IaC framework based on Python, Terraform, and Terragrunt managing AWS resources; championing IaC best practices while ensuring adherence to best practices for security, reliability, performance, cost optimization, and operational excellence. Responsibilities also include the development of internal tooling, modules, and libraries used by technology teams to both implement new projects and maintain and enhance existing platforms with a focus on automation, resiliency, availability, scalability, and performance that meet business needs within appropriate cost constraints, primarily leveraging open-source technologies and frameworks.

Numbers & Facts

LocationNY (
Remote
)
Salary$205,000–$212,500 Per Year

Description

Site Reliability Engineer I (SRE I) - DevOps Focus

Job Summary

How you will impact the organization…

Reporting to the Director, Infrastructure Operations, the SRE I (Site Reliability Engineer I) with a strong DevOps focus is a key member of a team responsible for supporting the tooling, pipelines, frameworks, and other technologies that underpin the many platforms deployed within the company's infrastructure. This role blends software engineering principles with deep DevOps expertise to automate and streamline the entire software delivery lifecycle. Additionally, the SRE I role will be an expert in Infrastructure as Code (IaC), AWS Cloud infrastructure design and best practices, and CI/CD platforms and processes, driving the automation and optimization of our operations.

As SRE I, they will partner closely with Engineering and Information Security peers on developing infrastructure solutions that follow established best practices and design patterns. They will also contribute to the continued development of RFCs, standards, and frameworks for IaC, automation, and supporting tools. Responsibilities also include the development of internal tooling, modules, and libraries used by technology teams to both implement new projects and maintain and enhance existing platforms with a focus on automation, resiliency, availability, scalability, and performance that meet business needs within appropriate cost constraints, primarily leveraging open-source technologies and frameworks.

Key Responsibilities

  • Partner with Engineering and Information Security peers to develop infrastructure solutions
  • Contribute to the development of RFCs, standards, and frameworks for IaC, automation, and supporting tools
  • Develop and maintain internal tooling, modules, and libraries for technology teams
  • Champion a DevOps culture and practices
  • Provide expert full-stack support to software engineering teams (Java, Python, Node, Go, etc.)
  • Design, implement, manage, and optimize CI/CD pipelines using tools like GitHub Actions and AWS CodePipeline
  • Maintain deep expertise in GitHub
  • Provide expert DevOps-focused full-stack guidance and support to software engineering teams
  • Champion and implement DevOps best practices across teams
  • Participate in grooming and prioritizing development efforts in extending and supporting the IaC, tooling, and infrastructure support application platforms

Value You Will Deliver

Lead the design, development, and maintenance of secure, scalable, resilient, and cost-effective cloud infrastructure solutions on AWS through a DevOps approach, leveraging the existing IaC framework based on Python, Terraform, and Terragrunt managing AWS resources; championing IaC best practices while ensuring adherence to best practices for security, reliability, performance, cost optimization, and operational excellence.

Design, implement, manage, and optimize robust CI/CD pipelines using tools like GitHub Actions and AWS CodePipeline for both infrastructure and applications; maintain deep expertise in GitHub.

Design, develop, and implement new tooling, applications, and platforms to improve and upgrade the capabilities of the IaC and automation platforms, and support infrastructure.

Provide expert DevOps-focused full-stack guidance and support to software engineering teams (using common languages such as Java, Python, Node, Go, etc.) to integrate DevOps practices, automate builds/deployments, identify/resolve reliability/performance bottlenecks, and establish comprehensive documentation.

Requirements

  • A security clearance or the ability to obtain a security clearance is required.
  • 6+ years hands-on experience managing and automating UNIX/Linux system environments within a DevOps context.
  • 5+ years of experience in a DevOps Engineer or SRE role with a strong DevOps focus, emphasizing infrastructure automation, CI/CD pipeline development, and cloud services.
  • 4+ years of experience designing and implementing infrastructure as code within the AWS ecosphere using Terraform.
  • Mastery in DevOps discipline and processes, including building and managing CI/CD pipelines supporting Infrastructure as Code frameworks such as Terraform, and Continuous Delivery Tools such as AWS CodePipeline, GitHub Actions, Jenkins, Git, Artifactory, etc.
  • Expert-level proficiency with Terraform and Terragrunt for managing AWS infrastructure as code.
  • Deep expertise in GitHub, GitHub Actions for CI/CD, including design, troubleshooting, and support.
  • Strong experience with AWS Cloud services (e.g., EC2, S3, RDS, VPC, IAM, Lambda, EKS/ECS, CloudWatch) and infrastructure design best practices, applied within a DevOps model.
  • Mastery of observability, monitoring, metrics and alerting at scale across regionally and globally resilient and distributed platforms leveraging common open source frameworks such as Prometheus, Thanos, OpenTelemetry, Grafana, etc.
  • Experience providing DevOps-centric support for applications developed in Java, Python, and Angular, including build automation, deployment pipelines, and observability.
  • Expert level proficiency in at least one scripting language (e.g., Python, Bash, Perl) and one programming language (e.g., Java, Go, Node).
  • Mastery of Containerization (Docker), and strong familiarity with the container ecosystem, especially Amazon ECS.
  • Mastery in config automation tool sets such as AWS Config and/or SSM, Puppet, Ansible, Chef, etc.
  • Hands-on experience with APM tools such as Zipkin, Jaeger, OpenTelemetry, NewRelic, etc.
  • Proficient with Jira, Confluence, and git toolset.
  • Hands-on experience with Agile/Scrum & Waterfall process environments.
  • Experience implementing and supporting a variety of Open Source frameworks and projects relevant to DevOps and SRE.

What You Will Stand Out If You Also Have

  • Experience with PCI, HiTrust, FedRamp/GovCloud and/or similar certification methodologies.
  • Experience with migrating and educating teams to newer SDLC and DevOps concepts.
  • Experience with APM/Observability and advanced DevOps/SRE concepts and methodologies.
  • Proven experience mentoring team members in DevOps practices.
  • Active, transferable U.S. Security clearance at the Public Trust level or higher preferred.

Additional Details

  • Location requirements: Fully remote within the U.S. (Los Angeles or Las Vegas preferred.)
  • Travel requirements: Up to 10% travel.
  • Physical requirements: Must be able to sit for long periods, as well as, handle long periods of screen time.
  • Technology requirements: Reliable, high speed internet.
  • Eligible for sponsorship: No
  • Security clearance: Ability to achieve and maintain a security clearance with the U.S. Government is required.
  • Salary range: $205,000 - $212,500 USD annually. The total compensation package for this position is negotiable and may also include annual performance bonus, ESPP, enhanced time off packages and benefits.

Similar Jobs