Title : Jr. GCP DevOps Engineer
Location: Dallas, TX & Portland, OR
Project Duration: 12 Months Contract
Job Type: Contract
Work Arrangement: Work from Client Office – 5 Days a Week
Interview: for Portland location Please apply if you can attend an in-person interview in Richardson, TX After 1st Virtual Interview.
Need Locals Only to Portland OR
Job Description
We are seeking a highly experienced GCP DevOps Engineer / Cloud Platform Engineer with strong production expertise in Google Cloud Platform, Kubernetes, GitOps, CI/CD, Infrastructure as Code, automation, security, and observability. The ideal candidate will have 7+ years of experience designing and supporting enterprise-scale DevOps/SRE environments and hands-on experience with GCP, GKE, GitHub Actions, ArgoCD, Terraform, Docker, and cloud-native technologies.
Key Responsibilities
DevOps Strategy & CI/CD
- Lead CI/CD best practices including Continuous Integration, Continuous Delivery, automated testing, TDD, and continuous improvement across infrastructure, application deployments, and incident response.
- Design, build, and scale enterprise GitOps delivery pipelines and cloud-native infrastructure across Google Cloud Platform (GCP).
- Build modular GitHub Actions workflows integrated with Google Artifact Registry, GCP-hosted infrastructure, Workload Identity Federation, and automated DevSecOps security gates.
- Implement and manage ArgoCD on Google Kubernetes Engine (GKE) for continuous delivery, drift reconciliation, automated synchronization, and progressive deployments across multi-cluster GCP environments.
- Use Helm and Kustomize to manage Kubernetes application configurations and environment-specific deployments.
- Automate GCP infrastructure using Terraform, enforcing cloud governance through IAM, Workload Identity, Secret Manager, Cloud KMS, Organization Policies, and Cloud Logging/Monitoring.
- Design resilient, scalable, highly available, and secure cloud infrastructure supporting enterprise and telecom-grade workloads.
- Build end-to-end deployment pipelines integrating GitHub, GitHub Actions, GKE, Google Artifact Registry, Cloud SQL, and other GCP services with automated promotion across Development, Staging, and Production.
- Champion Infrastructure as Code using Terraform and Ansible and establish standards for pipeline governance, security scanning, dependency management, and automated compliance.
Automation & Release Engineering
- Install, configure, and maintain GitHub Actions self-hosted runners on Linux and Windows environments with high availability and operational reliability.
- Develop production-grade Python, Bash, PowerShell, and Groovy automation scripts for deployment, backup, rollback, health checks, validation, and operational tasks.
- Automate application and database deployments across GCP environments.
- Implement automated release, rollback, validation, and health-check mechanisms.
- Integrate GitHub Actions with Jira for automated ticket lifecycle management, change tracking, approvals, and deployment traceability.
- Develop reusable CI/CD templates and automation frameworks for engineering teams.
GCP Cloud, Kubernetes & Infrastructure
- Design and implement cloud-native solutions using Google Cloud Platform, with emphasis on reliability, scalability, performance, and cost optimization.
- Hands-on administration and troubleshooting of Google Kubernetes Engine (GKE) clusters.
- Manage GCP services including:
- Google Compute Engine (GCE)
- Google Kubernetes Engine (GKE)
- Google Cloud Storage (GCS)
- Google Artifact Registry
- Cloud SQL
- Cloud Load Balancing
- VPC / Shared VPC
- Cloud NAT
- Cloud DNS
- IAM
- Secret Manager
- Cloud KMS
- Cloud Logging and Cloud Monitoring
- Pub/Sub
- Implement Kubernetes networking, RBAC, secrets management, ingress, autoscaling, service discovery, and workload security.
- Implement GCP cost optimization strategies across compute, storage, networking, and Kubernetes workloads.
Monitoring, Observability & SRE
- Implement enterprise monitoring and observability using Google Cloud Monitoring, Cloud Logging, Prometheus, Grafana, OpenTelemetry, Splunk, Zabbix, and Dynatrace.
- Define and maintain SLIs, SLOs, and SLAs for critical applications and infrastructure.
- Configure centralized logging, metrics, distributed tracing, dashboards, and alerting.
- Integrate monitoring with Slack and PagerDuty for real-time alerting and incident response.
- Participate in on-call rotations and troubleshoot critical production incidents.
- Lead post-incident reviews, root-cause analysis (RCA), and corrective/preventive action planning.
- Continuously improve platform reliability, deployment performance, and operational efficiency.
Security, Disaster Recovery & Operational Excellence
- Implement GCP security best practices including:
- IAM and least-privilege access
- Workload Identity
- Secret Manager
- Cloud KMS
- VPC Service Controls
- Organization Policies
- Branch protection
- GitHub Secrets
- RBAC
- Signed commits
- Implement automated security scanning and compliance checks within CI/CD pipelines.
- Design automated backup, disaster recovery, rollback, and business continuity strategies.
- Support environments requiring RTO <15 minutes and RPO <5 minutes.
- Conduct regular disaster recovery and failover exercises.
- Optimize deployment processes with a goal of reducing deployment time from approximately 30 minutes to less than 5 minutes.
- Improve infrastructure utilization and storage efficiency.
- Maintain detailed technical documentation, runbooks, architecture diagrams, and operational procedures.
- Mentor junior DevOps, Cloud, and SRE engineers.
Required Qualifications
- Bachelor's degree in Computer Science, Engineering, Information Technology, or related field, or equivalent experience.
- 7+ years of experience in DevOps, SRE, Infrastructure Automation, Cloud Engineering, or Platform Engineering.
- Strong production-level expertise in Google Cloud Platform (GCP), GKE, ArgoCD, GitHub Actions, and Terraform.
- Proven experience designing and implementing enterprise-scale CI/CD pipelines.
- Strong hands-on experience with GKE / Kubernetes, Docker, Helm, Kustomize, and GitOps.
- Strong experience with GitHub Actions and CI/CD automation.
- Strong Infrastructure as Code experience using Terraform.
- Strong scripting/programming skills in Python, Bash, PowerShell, and/or Groovy.
- Experience with Ansible, Puppet, or Chef.
- Hands-on experience with:
- GCP
- GKE
- Google Artifact Registry
- Cloud SQL / database deployment automation
- IAM
- VPC
- Cloud Storage
- Secret Manager
- Cloud Monitoring
- Cloud Logging
- Docker
- Kubernetes
- Experience with AWS or Azure is a plus.
- Strong knowledge of monitoring, observability, and logging platforms.
- Experience integrating DevOps platforms with Jira, Slack, and PagerDuty.
- Experience supporting high-availability and mission-critical production environments.
- Strong troubleshooting, analytical, communication, and stakeholder-management skills.
- Experience with cloud security, compliance, disaster recovery, and production incident management.