Wednesday, September 16, 2026
CLOUDSUTRA

Zensar Hiring: Site Reliability Engineer (SRE)

ZensarHyderabad, Telangana, India (Work from Anywhere Option Available)
INR Competitive/year
Zensar
DevOpsFull-timePosted 27 April 2025Updated 27 April 2025Expired
Share:

Job Description


Zensar is looking for a talented Site Reliability Engineer (SRE) to design, implement, and maintain highly available and scalable systems. As an SRE, you will play a key role in ensuring system reliability, performance, and continuous improvement of critical platforms.

Requirements


  • Strong knowledge of Linux/Unix systems and command-line tools.
  • Proficiency in scripting languages: Python, Shell, or Perl.
  • Experience with configuration management tools: Ansible, Puppet, or Chef.
  • Familiarity with cloud platforms: AWS, Azure, or Google Cloud.
  • Understanding of networking protocols: TCP/IP, HTTP, DNS, etc.
  • Knowledge of containerization tools: Docker, Kubernetes.
  • Familiarity with monitoring and logging tools: Prometheus, Grafana, ELK Stack, Splunk (good to know).
  • Experience with Citrix technologies like XenApp, XenDesktop, NetScaler (Optional but valuable).
  • Understanding of virtualization: VMware, Hyper-V.
  • Basic skills in Terraform syntax and GitLab CI/CD pipeline configurations.
  • Experience provisioning and configuring cloud resources via CLI/API.
  • Understanding of basic logging queries and operational troubleshooting in Linux systems.

Secondary Skills Required

  • Bachelor’s degree in Computer Science, Engineering, or related field.
  • Proven experience as an SRE or similar role.
  • Solid understanding of DevOps principles and Agile methodologies.
  • Familiarity with CI/CD pipelines and source control tools (Git, SVN).
  • Knowledge of security best practices in production environments.
  • Relevant certifications are a plus (e.g., AWS DevOps Engineer, CKA).

Role & Responsibilities

  • Design and maintain highly available, scalable systems.
  • Define and establish SLOs and SLAs for critical applications.
  • Proactively monitor system health and resolve performance issues.
  • Build tools, dashboards, and alerts for better visibility and control.
  • Conduct post-incident reviews and implement preventive measures.
  • Automate routine tasks and optimize operational processes.
  • Document architecture, configurations, and troubleshooting procedures.
  • Support capacity planning, resource allocation, and system scaling.
  • Collaborate with development teams to meet performance standards.
  • Stay updated with industry best practices and emerging SRE trends.

Objectives of This Role

  • Manage the production environment and monitor overall system health.
  • Build and maintain infrastructure and platform management software.
  • Improve reliability, quality, and time-to-market for solutions.
  • Optimize performance and drive innovations in system capabilities.
  • Provide operational support and engineering for large-scale distributed systems.

Work Environment

  • Flexible work options: Work from Hyderabad office or remotely.
  • Dynamic Rotating Shifts to cover critical operations globally.
  • Culture of continuous improvement, innovation, and collaboration.

#SREJobs #LinuxEngineer #CloudEngineer #DevOpsEngineer #SiteReliabilityEngineer #HyderabadJobs #RemoteJobs #CloudInfrastructure #Docker #Kubernetes #Terraform #Ansible #ZensarCareers