Site Reliability Engineering, Staff

Job Description
The Engineering Excellence Group at Synopsys focuses on accelerating innovation and automating enterprise infrastructure.
The team is responsible for enhancing productivity, agility, and reliability across Synopsys products and IT operations.
As an SRE (Site Reliability Engineer), Staff, you will play a key role in improving infrastructure reliability, performance, and automation while collaborating with engineering and business teams.
Requirements
Key Roles & Responsibilities
πΉ Infrastructure Optimization:
-
Design and implement IT infrastructure improvements for enhanced reliability and performance.
-
Drive standardization and automation of infrastructure components.
πΉ Collaboration & Enhancements:
-
Work with engineering teams and business units to integrate SRE best practices.
-
Translate business and technical requirements into scalable infrastructure solutions.
πΉ Automation & Monitoring:
-
Implement monitoring systems with intelligent alerts for quick issue resolution.
-
Continuously improve resource utilization with automation and internal tools.
πΉ Security & Incident Management:
-
Maintain vulnerability management policies using a risk-based priority approach.
-
Work with security teams to mitigate risks and improve infrastructure security.
πΉ Troubleshooting & On-Call Support:
-
Conduct root cause analysis for production issues and implement fixes.
-
Participate in off-hours maintenance and an on-call rotation schedule.
Required Skills
π₯ Infrastructure & Cloud Computing:
β Strong experience with Linux, Windows, high-performance computing, and storage platforms.
β Expertise in cloud services (AWS, GCP, Azure), containerization (Docker, Kubernetes, OpenStack, ECS, Mesos), and virtualization.
β Understanding of networking, service dependencies, and troubleshooting.
π» Development & Automation:
β Hands-on experience with infrastructure automation tools: Terraform, Ansible, Jenkins, GitHub.
β Proficiency in programming languages: Python, Java, Go, Node.js.
β Experience in CI/CD pipelines, SDLC, and Agile development.
π Security & Monitoring:
β Strong knowledge of security hardening (CIS benchmarks), SIEM tools, and ITIL processes.
β Experience in incident management, service mapping, and CMDB (ServiceNow).
π Problem-Solving & Collaboration:
β Passion for debugging distributed systems and finding technical root causes.
β Strong communication skills and ability to work across teams.
Experience & Education
π Bachelorβs/Masterβs degree in Computer Science, IT, or a related field.
π 8+ years of experience in IT infrastructure & operations, with 4+ years in an SRE role.
π Certification in AWS, ITIL, or Kubernetes is a plus.
Why Join Synopsys?
π Innovative Work β Be part of cutting-edge infrastructure automation.
π Global Impact β Help optimize mission-critical systems.
π Career Growth β Work with top engineering teams and learn new technologies.
π Diversity & Inclusion β Synopsys is an equal opportunity employer.
If you are passionate about automation, infrastructure, and reliability engineering, apply now and be part of Synopsys' Engineering Excellence Group! π