Posted 25 July, 2026
Principal Site Reliability Engineer (Linux/Networking/Automation)
Zscaler
Hyderabad, IND
Full Time
Reference: 102_705768_5089209007
Role
We are looking for a Principal Site Reliability Engineer to join our team. This is a hybrid role based in Hyderabad, reporting to the Director, Site Reliability Engineering in the Engineering department. You will contribute as a development engineer within our Engineering team, helping to build and enhance the world's largest cloud security platform. You will bring your vision and passion to a team of experts enabling organizations worldwide to harness speed and agility through a cloud-first strategy.
What you'll do (Role Expectations)
- Perform operational duties for FedRAMP cloud products, including deployments, on-call support, and incident management
- Join deployment sync calls and conduct Operations hand-offs to ensure seamless continuity
- Manage cloud infrastructure elements, including AWS GovCloud, private cloud, containers, and VMs
- Operate and enhance monitoring systems while driving automation, scripting, and Infrastructure as Code (IaC) efforts
- Write and maintain documentation, resolve escalations, prevent incident recurrence, and promote DevOps best practices
Who You Are (Success Profile)
- You thrive in ambiguity. You're comfortable building the path as you walk it. You thrive in a dynamic environment, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful.
- You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. True ownership involves leveraging dynamic range: the ability to navigate seamlessly between high-level strategy and hands-on execution.
- You are a problem-solver. You love running towards the challenges because you are laser-focused on finding the solution, knowing that solving the hard problems delivers the biggest impact.
- You are a high-trust collaborator. You are ambitious for the team, not just yourself. You embrace our challenge culture by giving and receiving ongoing feedback-knowing that candor delivered with clarity and respect is the truest form of teamwork and the fastest way to earn trust.
- You are a learner. You have a true growth mindset and are obsessed with your own development, actively seeking feedback to become a better partner and a stronger teammate. You love what you do and you do it with purpose.
What We're Looking for (Minimum Qualifications)
- Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain
- 10+ years of experience as a Site Reliability Engineer with expertise in Operations and Engineering
- Experience with FedRAMP compliance (High/Moderate levels), vulnerability management, and continuous monitoring, including scanning, patching, and reporting
- Proficiency in Linux administration, network troubleshooting, and infrastructure as code (Ansible, Terraform) in cloud environments
- Experience in large-scale distributed systems, containerized architectures (AWS ECS, Kubernetes), and cloud services, with a strong foundation in web security, networking, and coding (Python)
What Will Make You Stand Out (Preferred Qualifications)
- Experience implementing AIOps frameworks, leveraging machine learning for predictive cloud infrastructure autoscaling, or utilizing AI-driven log anomaly detection tools to optimize root-cause analysis
- Experience with containerized architectures such as AWS ECS and Kubernetes
- Knowledge of web security protocols including HTTP, SSL/TLS, DNS, SQL, and networking fundamentals
#LI-SK3
#LI-HYBRID