NOC Analyst - Cloud Operations - (24/7 Environment)
NOC Analyst
Location: Bangalore (24 / 7 Environment)
Role Summary
We are looking for a Cloud Operations & NOC Analyst who will act as the first line of operational defense for enterprise infrastructure, cloud platforms, and applications. This role requires strong real-time monitoring, incident response, and troubleshooting capabilities, along with a proactive mindset toward improving operational processes and reducing alert noise.
Key Responsibilities
Monitoring & Incident Management
- Monitor infrastructure, applications, and cloud platforms using tools such as New Relic, Datadog, Prometheus/Grafana, AWS CloudWatch, or GCP Monitoring
- Perform real-time alert triage, validation, and troubleshooting to restore services quickly
- Act as the first responder for incidents, ensuring minimal downtime and impact
- Identify false positives and reduce alert noise through analysis and tuning
Incident Handling & Escalation
- Own and manage high-priority incidents (P1/P2), including:
- Driving incident bridges
- Coordinating with cross-functional teams (SRE, CloudOps, Network, Security)
- Tracking progress until resolution
- Escalate complex issues with clear documentation and impact analysis
Operations & Process Execution
- Execute runbooks and SOPs for standard incidents and operational tasks
- Contribute to creating and improving runbooks based on recurring issues
- Perform routine health checks on servers, networks, and failover systems
Maintenance & Reliability
- Monitor systems during maintenance windows and validate post-change stability
- Ensure redundancy mechanisms (failover, load balancers) are functioning
- Monitor backups, batch jobs, and scheduled processes, resolving failures proactively
Documentation & Reporting
- Maintain accurate incident records in ticketing systems (Jira/ServiceNow)
- Contribute to daily/weekly operational health and performance reports
- Support RCA documentation and post-incident reviews
Security Awareness
- Identify potential security-related anomalies and escalate to SOC/security teams
Required Skills
Technical Skills
- Experience with monitoring & alerting tools:
- New Relic / Datadog / Prometheus / Grafana
- Cloud monitoring (AWS / GCP)
- PagerDuty or similar alerting systems
- Basic knowledge of cloud environments and application availability & performance
- Familiarity with firewalls and security principles
- Good understanding of On Prem, Cloud Infra, VMs, EC2 Instances & EKS Clusters
- Basic understanding of networking fundamentals & protocols
Operational Skills
- Strong troubleshooting and analytical thinking
- Experience in incident management and escalation workflows
- Ability to work in a high-pressure, 24/7 operations environment
- Good documentation and communication skills
Experience & Qualifications
- 4 - 5 years in NOC / Infrastructure Monitoring / Cloud Operations / SOC
- Bachelor's degree in IT / Computer Science or related field
- Experience working in enterprise or SaaS environments preferred
Good-to-Have Certifications
- AWS Certified Cloud Practitioner
- Google Associate Cloud Engineer
- CompTIA Cloud+
- ITIL V4 or V5
- CCNA
#LI-DS9
#LI-Bengaluru
#LI-Hybrid
#LI-CloudOperations