Skip to main content
Posted 22 July, 2026

Azure SRE lead engineers

Diverse Lynx
bengaluru,Karnataka,560063 Full Time
Reference: 365_569689_26-01942

Description:

We are seeking a Site Reliability Engineer (SRE) with 7+ years of experience to support and enhance the reliability, availability, and performance of critical banking systems at Truist. The role requires strong handson expertise in cloudnative platforms, observability, automation, and incident management, with a focus on reliability engineering and operational excellence

Required Technical Skills

Cloud & Infrastructure

  • Microsoft Azure
  • Kubernetes
  • OpenShift

Observability & Monitoring

  • Datadog
  • Dynatrace / AppDynamics
  • Splunk
  • Jenkins
  • Ansible

Automation & CI/CD

  • Python
  • Kafka
  • RabbitMQ
  • Exposure to Java and Node.js

Production Support & Incident Management

  • Strong experience handling major incidents
  • Production support in highavailability, missioncritical environments
  • Root cause analysis and reliability improvement
  • Engineer and enhance observability across systems and platforms
  • Define, implement, and track SLIs and SLOs
  • Design and build automation for recovery and selfhealing
  • Apply cloudnative resiliency and failureisolation patterns
  • Lead major incident response with an engineeringdriven approach
  • Drive systemlevel root cause fixes
  • Reduce longterm incident volume through reliability engineering initiatives

Desired Skills

  • Analyze, optimize, and enable CI/CD pipelines to improve reliability outcomes

Supplementary Skills (Good to Have)

Advanced use of AIOps for predictive reliability insights

Sign up for Job Alerts