Site Reliability Engineer (SRE) – Google Cloud Platform (GCP)
Location: Chandler, AZ (Hybrid)
Duration: 12+ Month Contract
***This is a W2 Opportunity ONLY!!!- ABSOLUTELY NO C2C
Job Overview
We are seeking a Site Reliability Engineer (SRE) with strong Google Cloud Platform (GCP) and Terraform expertise to support the design, automation, and reliability of enterprise cloud infrastructure. This role will focus on building and maintaining scalable, secure, and highly available cloud environments using Infrastructure as Code (IaC), while driving automation, observability, and operational excellence. The ideal candidate has experience with Terraform Enterprise, CI/CD, DevSecOps, and cloud platform engineering in large-scale environments.
Responsibilities
- Design, build, and maintain Google Cloud Platform (GCP) infrastructure using Terraform and Infrastructure as Code (IaC).
- Develop reusable Terraform modules and standardized infrastructure patterns for consistent cloud provisioning.
- Implement and maintain scalable, secure, and compliant cloud environments aligned with enterprise best practices.
- Build and enhance CI/CD pipelines for automated infrastructure deployment, testing, and validation.
- Support Terraform Enterprise, including automated provisioning, policy enforcement, and infrastructure governance.
- Implement automation for provisioning, configuration management, and cloud platform operations.
- Integrate security and compliance controls into infrastructure delivery through DevSecOps and policy-as-code practices.
- Design and improve observability solutions, including monitoring, logging, alerting, dashboards, and distributed tracing.
- Define and support Service Level Objectives (SLOs), availability, latency, and performance metrics.
- Participate in incident response, troubleshooting, root cause analysis, and continuous service improvement.
- Optimize cloud infrastructure for scalability, performance, reliability, and cost efficiency.
- Conduct capacity planning and performance testing to ensure resilient cloud services.
- Collaborate with engineering, architecture, and security teams to deliver reliable and secure cloud platforms.
- Continuously identify opportunities to automate manual processes and improve operational efficiency.
- Evaluate emerging cloud technologies and automation tools, including AI/ML capabilities where applicable.
Required Qualifications
- 4–8+ years of experience in Site Reliability Engineering, Cloud Infrastructure Engineering, Platform Engineering, or Cloud Operations.
- Hands-on experience with Google Cloud Platform (GCP).
- Strong experience with Terraform and Infrastructure as Code (IaC); Terraform Enterprise experience is highly preferred.
- Experience developing reusable Terraform modules and managing infrastructure through code.
- Experience building and maintaining CI/CD pipelines for infrastructure deployments.
- Knowledge of DevSecOps practices and integrating security into automated deployment workflows.
- Strong understanding of GCP services, including VPC networking, IAM, load balancing, and cloud architecture fundamentals.
- Experience with policy-as-code, governance, and cloud compliance frameworks.
- Experience with monitoring, logging, and observability tools.
- Hands-on experience with incident management, troubleshooting, and root cause analysis in cloud environments.
- Experience designing dashboards, alerts, and monitoring strategies aligned with SLOs.
- Strong scripting and automation skills.
- Excellent problem-solving, analytical, and communication skills.
- Ability to work effectively within cross-functional Agile teams.
Preferred Qualifications
- Experience with Terraform Enterprise administration.
- Experience supporting large-scale enterprise cloud environments.
- Familiarity with AI/ML-driven operational automation and observability.
- Experience implementing cloud governance and compliance standards.
- Experience with cloud-native monitoring and distributed systems.