ResponsibilitiesDesign, implement, and manage scalable, reliable infrastructure on GCPMaintain and improve system availability, performance, and latencyBuild and manage infrastructure as code using tools like Terraform or similarDevelop automation to reduce manual operational workMonitor systems using observability tools (metrics, logs, tracing) and respond to incidentsParticipate in on-call rotations and lead incident response and postmortemsCollaborate with development teams to improve service reliability and deployment processesOptimize cloud resource usage and cost efficiencyImplement and maintain CI/CD pipelines for reliable software deliveryDefine and track SLIs, SLOs, and error budgetsRequired Knowledge, Skills, And Abilities3+ years of experience in Site Reliability Engineering, DevOps, or similar rolesStrong experience with Google Cloud Platform (GCP) services (e.g., Compute Engine, GKE, Cloud Run, Cloud Storage)Experience with containerization and orchestration (Docker, Kubernetes)Proficiency in at least one programming language (e.g., Python, Go, Java)Experience with Infrastructure as Code (Terraform preferred)Solid understanding of networking, security, and distributed systemsExperience with monitoring and logging tools (e.g., Prometheus, Grafana, Cloud Monitoring)Familiarity with CI/CD pipelines and automation toolsHands-on experience with Google Cloud Platform, including:GCP: GKE, Compute Engine, Cloud Storage, Pub/Sub (or equivalents)Cloud Monitoring & LoggingBigQueryDataflowDatastreamIAM and networkingComposer/AirflowKubernetes: deployment, scaling, reliability patternsCI/CD: GitHub Actions, GitLab CI, or similarObservability: GCP Cloud Monitoring, LoggingExperience operating systems in 24/7 production environmentsMinimum QualificationsBachelor's degree in Business, Information Technology, Computer Science, or a related field.3+ years experience in Site Reliability Engineering, Cloud Platform Engineering, or DevOps3+ years operating production workloads on Google Cloud Platform (GCP)Ability to understand and speak English at a level of proficiency allowing employee to issue, receive and respond to both safety and operations-related directions in EnglishPreferred QualificationsOil and Gas Industry knowledgeTechnology/Digital Industry knowledgeAbout UsThe Evolving Oil Field Demands Evolving Service ProvidersNexTier is a leading provider of integrated completions that employs sustainable practices and equipment to support our customers' ESG goals while accelerating production in the most demanding US land basins.#J-18808-Ljbffr. Job DescriptionThe NexTier Technology team is looking for a Site Reliability Engineer (SRE) to help build, scale, and maintain highly reliable systems on Google Cloud Platform (GCP).