Senior SRE Engineer with Azure

Bangalore,Trivandrum,Kochi,,Chennai, · Onsite Full-time 5-7yrs B.Tech/MCA Direct Client hiring Posted 5 months ago💰 Not disclosed
Apply now

Must Have Skills

AzureAny monitoring toolPython scriptingTerraformKubernetesPythonCloud InfrastructureDevOpsCI/CDMonitoring ToolsNIL

Preferred Skills

AzureAny monitoring toolPython scriptingTerraformKubernetesPythonCloud InfrastructureDevOpsCI/CDMonitoring ToolsNIL

Job Description – Azure Site Reliability Engineer (SRE)

Role: Azure Site Reliability Engineer (SRE)

Experience: 5–7

Location: Kochi

Employment Type: Full-Time

Role Overview

We are seeking a highly skilled Azure Site Reliability Engineer (SRE) to join our Cloud Engineering team. The ideal candidate will have strong expertise in Terraform and Python scripting, with proven experience designing, automating, and operating highly available, scalable, and secure infrastructure on Microsoft Azure.

This role requires a proactive engineer passionate about automation, reliability, performance optimization, and operational excellence in cloud environments.

Key Responsibilities

Cloud Infrastructure & Automation

  1. Architect, deploy, and manage secure, scalable, and highly available infrastructure on Microsoft Azure.
  2. Develop and maintain Infrastructure as Code (IaC) using Terraform.
  3. Build automation tools and frameworks using Python and Shell scripting.
  4. Establish standardized and reusable deployment patterns across cloud environments.

CI/CD & DevOps Enablement

  1. Design and implement CI/CD pipelines for infrastructure and application deployments.
  2. Promote DevOps best practices, including automated testing and release management.
  3. Enable seamless environment provisioning and configuration management.

Reliability & Monitoring

  1. Implement monitoring, alerting, and observability solutions using Azure-native tools.
  2. Continuously monitor system performance, availability, and reliability.
  3. Perform root cause analysis (RCA) and implement preventive solutions to minimize incidents.

Performance & Cost Optimization

  1. Optimize cloud resource utilization to improve performance and scalability.
  2. Drive cost optimization initiatives across Azure environments.
  3. Automate operational workflows to reduce manual effort and operational toil.

Collaboration & Engineering Excellence

  1. Partner with development teams to improve application resilience and fault tolerance.
  2. Contribute to incident management, postmortems, and capacity planning.
  3. Advocate for reliability-first engineering practices and operational excellence.

Required Skills & Qualifications

  1. Strong hands-on experience with Microsoft Azure cloud services.
  2. Proven expertise in Terraform (Infrastructure as Code).
  3. Strong scripting skills in Python (Shell scripting is a plus).
  4. Experience implementing CI/CD pipelines and DevOps workflows.
  5. Solid understanding of cloud networking, security, and monitoring concepts.
  6. Experience in production support and reliability engineering practices.

Key Skills

Azure | Site Reliability Engineering (SRE) | Terraform | Python | Monitoring Tools | DevOps | CI/CD | Cloud Infrastructure

Apply to this job

Required — upload a file or paste the text