8+ Years Role Overview We are seeking a Cloud Site Reliability Engineer (SRE) to support and enhance our Azure cloud platform, infrastructure operations, and observability ecosystem. This role will focus on infrastructure automation, platform reliability, monitoring, deployment readiness, and cloud modernization initiatives. The ideal candidate will bring strong Azure expertise, hands-on Terraform experience, and a proactive approach to operational excellence. Key Responsibilities Support and maintain Azure-based applications and cloud infrastructure. Manage cloud environments, infrastructure provisioning, and deployment pipelines. Design and implement monitoring, observability, and application health solutions. Drive platform modernization initiatives, including Azure Container Apps (ACA) adoption. Partner with development teams to improve deployment readiness, reliability, and operational efficiency. Contribute to cloud architecture, platform engineering, and technology decision-making. Automate infrastructure and operational processes using Infrastructure-as-Code principles. Troubleshoot performance, availability, and scalability issues across cloud environments. Required Qualifications Terraform (Highest Priority) Strong hands-on experience with Terraform for Infrastructure-as-Code (IaC). Experience automating cloud infrastructure provisioning and management. Microsoft Azure Strong understanding of Azure services and cloud-native architecture. Experience managing Azure infrastructure and platform operations. Knowledge of security, scalability, and high-availability principles within Azure. Azure Container Apps (ACA) Experience deploying and managing containerized workloads. Familiarity with Azure Container Apps and modern application hosting strategies.
Cloud Networking Strong understanding of:
Azure Virtual Networks (VNets) Cloud networking concepts Connectivity and network architecture Hybrid and cloud-native networking solutions Monitoring & Observability Experience designing and implementing observability solutions. Strong focus on application performance, reliability, and operational visibility. Grafana OpenTelemetry Application Performance Monitoring (APM) tools Logging and alerting platforms DevOps & Platform Engineering CI/CD pipeline implementation Containerization technologies Deployment automation Platform reliability and operations Nice-to-Have Qualifications Advanced Azure architecture experience. Platform Engineering or DevOps Engineering background. Experience building large-scale cloud observability solutions. Expertise in infrastructure automation and cloud modernization initiatives. Ideal Candidate Profile 5+ years of experience in Cloud Engineering, SRE, DevOps, or Platform Engineering. Strong Azure and Terraform expertise. Monitoring and observability focused mindset. Comfortable working across infrastructure, applications, and platform operations. Collaborative, adaptable, and eager to learn new technologies. Practical problem solver with a broad cloud engineering skill set rather than a narrowly focused infrastructure specialization.