SRE/DevOps Engineer – Kubernetes | ServiceNow | Hyderabad

Job Description
Supercharge Global Infrastructure with Kubernetes and Observability – Join ServiceNow
ServiceNow, the intelligent cloud platform powering over 8,000 enterprises globally, is hiring an SRE/DevOps Engineer with Kubernetes mastery. As a key member of our AI Services Team, you will develop and scale cloud-native systems for high availability, observability, and performance. This role is ideal for engineers who are passionate about Kubernetes, monitoring, SLAs, and driving reliability-focused engineering culture.
Requirements
- Administer, optimize, and troubleshoot Kubernetes clusters and container orchestration
- Build robust cloud-based AI/ML platforms used by Fortune 500 companies
- Set up and manage monitoring and observability tools like Prometheus, Grafana, ELK Stack
- Implement OpenTelemetry to unify telemetry across services and boost reliability
- Define and track SLIs, SLOs, and SLAs, ensuring high availability and minimal latency
- Automate workflows using Ansible, Puppet, and integrate with GitOps/CD pipelines
- Participate in on-call rotations, incident response, and post-mortem reviews
- Collaborate with engineering, security, and product teams to drive DevOps excellence
- Apply cloud security best practices including encryption and IAM policies
- Contribute to disaster recovery, cost optimization, and cloud governance efforts
Core Skills & Experience
- 2+ years of SRE/DevOps experience with Kubernetes in production environments
- Deep expertise in monitoring, logging, and observability practices
- Proficient in Java or Python, with strong understanding of OOP and system design
- Experience with public cloud platforms – Azure preferred, AWS optional
- Strong Linux/Unix system administration and scripting skills
- Infrastructure-as-Code proficiency using Ansible, Puppet, or Terraform
- Clear grasp of SLI/SLO/SLA concepts and error budgets
- Effective communicator and proactive learner in dynamic engineering environments
Bonus/Preferred Skills
- GitOps and CI/CD with Jenkins, ArgoCD, Git, Rundeck
- Database skills in MySQL and Hadoop
- Cloud cost management strategies and billing optimization
- Familiarity with ITIL processes and business continuity planning
- Awareness of data encryption, access control, and cloud identity management
- Change advocate mindset: able to influence, improve, and drive reliability practices
Why ServiceNow?
- Be part of an elite team shaping the future of AI-powered enterprise platforms
- Build solutions that serve 85% of the Fortune 500
- Enjoy a flexible hybrid work culture with real growth potential
- Work with a product-first mindset in a fast-paced, scalable tech environment
#KubernetesJobs #SREIndia #DevOpsCareers #Observability #PrometheusMonitoring #GrafanaDashboards #OpenTelemetry #SLISLO #CloudSecurity #CostOptimization #MySQLJobs #HadoopDevOps #CloudReliability #AzureJobs #GitOps #ArgoCD #AnsibleEngineer #PythonDevOps #HybridJobsIndia #ServiceNowJobs