Seeking a skilled Platform Engineering SRE with expertise in Azure and Azure Kubernetes Service (AKS) to build and operate highly reliable, scalable, and secure cloud platforms. This role combines software engineering principles with infrastructure expertise to ensure platform resilience, automation, and operational excellence.
Key Responsibilities:
Platform Engineering & Azure Infrastructure:
Design, build, and maintain cloud-native platforms on Microsoft Azure
Architect and manage Azure Kubernetes Service (AKS) clusters for production workloads
Implement Infrastructure as Code (IaC) using Terraform
Design multi-region, highly available architectures leveraging Azure services
Kubernetes & Container Orchestration (AKS Focus):
Deploy, manage, and optimize AKS clusters
Ensure Kubernetes best practices (RBAC, network policies, pod security, etc.)
Manage containerized workloads using Docker
Troubleshoot cluster performance, networking, and workload issues
Reliability & Operations:
Lead incident response, production support, and post-incident reviews
Implement observability for underlying systems
Ensure high availability, minimal downtime and scalability for platform services
Automation & DevOps:
Build and maintain CI/CD pipelines
Automate infrastructure provisioning, deployments, and operational runbooks
Improve developer experience through self-service platform capabilities
Security & Governance:
Implement Azure security best practices including Managed Identities, Key Vault, and RBAC
Secure Kubernetes clusters (network policies, pod identity, secrets management)
Ensure compliance with enterprise security and governance standards
Required Qualifications:
3+ years experience in SRE / DevOps / Platform Engineering
Strong hands-on experience with Microsoft Azure
Expertise in Azure Kubernetes Service (AKS)
Experience with Docker and Kubernetes ecosystem tools
Proficiency in Infrastructure as Code (Terraform)
Experience with CI/CD tools (Azure DevOps preferred)
Strong scripting/programming skills (Python, Go, Bash)
Solid understanding of Linux, networking, and distributed systems
Key Skills:
AKS cluster administration and troubleshooting
Cloud-native architecture design on Azure
Automation-first mindset
Observability and monitoring expertise
Incident management and root cause analysis
Cross-team collaboration and communication
The pay range that the employer in good faith reasonably expects to pay for this position is $36.98/hour - $57.79/hour. Our benefits include medical, dental, vision and retirement benefits. Applications will be accepted on an ongoing basis.
Tundra Technical Solutions is among North America’s leading providers of Staffing and Consulting Services. Our success and our clients’ success are built on a foundation of service excellence. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. Qualified applicants with arrest or conviction records will be considered for employment in accordance with applicable law, including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Unincorporated LA County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: client provided property, including hardware (both of which may include data) entrusted to you from theft, loss or damage; return all portable client computer hardware in your possession (including the data contained therein) upon completion of the assignment, and; maintain the confidentiality of client proprietary, confidential, or non-public information. In addition, job duties require access to secure and protected client information technology systems and related data security obligations.