Skip to main content
Posted 27 June, 2026

DevOps Engineer

NAVVYASA CONSULTING PRIVATE LIMITED
Gurugram, HR, IN Full Time
Reference: 6e9c6b909b81017b

Job Description

Role Overview:

\n Experienced DevOps engineer who can own and scale production infrastructure end-to-end - from CI/CD and IaC to observability and incident response. You’ll lead design docs, harden reliability and security, drive cost/perf efficiency.

\n

What You’ll Do

\n

%CF; Architect and maintain CI/CD pipelines (build, test, security scans, deploy, rollback) with quality gates and environment promotions.

\n

%CF; Design and operate container platforms (ECS/EKS or equivalent), service discovery, blue/green & canary strategies, and autoscaling.

\n

%CF; Implement Infrastructure as Code (Terraform/CDK/CloudFormation), enforce modular, reviewable, and drift-free infra.

\n

%CF; Build observability: metrics/logs/traces, SLOs/SLIs, dashboards, and actionable alerts; reduce MTTR through runbooks and automation.

\n

%CF; Champion platform reliability: capacity planning, HA/DR (multi-AZ), backup/restore testing, change management.

\n

%CF; Own secrets management, IAM least-privilege, network policies, and baseline hardening (CIS where relevant).

\n

%CF; Drive cost optimization (rightsizing, autoscaling policies, savings plans/spot, storage lifecycle) with monthly reporting.

\n

%CF; Establish release/incident processes (postmortems, RCAs) and lead remediation to cut change failure rate.

\n

%CF; Partner with Backend/AI/Frontend teams to productize models/services (GPU pools, batching, caching layers) and streamline developer workflows.

\n

%CF; lead design reviews, tech spikes, Monitoring and documentation.

\n

Technical Qualifications

\n

%CF; 2-3+ years in DevOps/SRE/Platform roles supporting production systems at scale.

\n

%CF; Strong with AWS : VPC, IAM, ECS/EKS, ALB/NLB, RDS/Elasticache/Object storage, CloudWatch.

\n

%CF; Proficient in Terraform (or CDK/CloudFormation), CI/CD (GitHub/GitLab/Jenkins/Argo) including artifacts and environment promotion.

\n

%CF; Containers & orchestration: Docker, task definitions/helm charts, autoscaling, health checks, readiness/liveness.

\n

%CF; Observability: Prometheus/Grafana, OpenTelemetry, log pipelines (ELK/CloudWatch/Datadog), alert routing.

\n

%CF; Networking & security: VPC/Subnets, SGs/NACLs, TLS, DNS, WAF, IAM design, secrets (KMS/Parameter Store/Vault).

\n

%CF; Scripting/automation in Python/Bash, configuration management (Ansible or equivalent).

\n

%CF; Proven incident management: on-call practice, runbooks, RCAs, tuning alerts to reduce noise.

\n

Nice to Have

\n

%CF; Kubernetes (EKS) production experience, service mesh (Istio/Linkerd), GitOps (ArgoCD/Flux).

\n

%CF; Image and dependency security (Trivy/Grype/Snyk), SBOMs, policy-as-code (OPA/Conftest).

\n

%CF; Data platform ops (Mysql/Postgres/PITR, replicas), streaming (Kafka/Kinesis).

\n

%CF; All the corresponding services in azure

\n

Startup-Specific Expectations

\n

%CF; Be comfortable with ambiguity and a fast-paced, evolving environment.

\n

%CF; Proactively take on varied technical tasks outside your comfort zone.

\n

%CF; Help reduce operational toil via automation and smarter tooling.

\n

%CF; Contribute ideas on performance, cost savings, and process improvements.

Sign up for Job Alerts