Site Reliability, Staff

Technology, Data & Digital · IT Infrastructure & Security · Site Reliability Engineering · DevOps · Software Engineering

In short

We are looking for an experienced Site Reliability, Staff engineer to lead a 24x7 Linux operations team in Noida, India. This role requires strong people management and technical leadership in Linux system administration, SRE, and automation. You will ensure operational excellence, manage incidents, and drive continuous improvement.

Responsibilities

  • Lead and manage a 24x7 L1 Linux Engineering / SRE team in rotational shifts.
  • Oversee hiring, onboarding, performance management, coaching, and career development for L1 engineers.
  • Manage L1 production support operations for Linux systems.
  • Serve as the first leadership escalation point during major production incidents.
  • Ensure adherence to SLAs, OLAs, and operational KPIs like availability and MTTR.
  • Provide technical oversight for Linux OS, bare metal and virtualized platforms, and monitoring/logging systems.
  • Drive automation adoption using Ansible, Bash, and Python.
  • Define and maintain SOPs, runbooks, escalation procedures, and documentation.
  • Partner with platform, network, security, and engineering teams to enhance system reliability.

Requirements

  • 10–14+ years of experience in IT Infrastructure, Linux Operations, or SRE.
  • 4–6+ years of people management experience, preferably managing 24x7 support teams.
  • Strong hands-on background in Linux system administration and production support.
  • Experience with incident management, on-call models, and rotational shifts.
  • Advanced knowledge of Linux OS internals.
  • Experience with virtualization platforms (VMware, KVM, OpenStack, oVirt).
  • Knowledge of monitoring and logging tools (e.g., Nagios, ELK).
  • Experience with automation and configuration management (Ansible).
  • Scripting skills in Bash and/or Python.

Desired Qualifications

  • A strong people leader with excellent coaching and decision making skills.
  • Calm and effective under high pressure production scenarios.
  • Highly structured and data driven in driving operational excellence.
  • An effective communicator and stakeholder partner.
  • Passionate about reliability engineering, automation, and continuous improvement.

Benefits

  • Opportunity to lead mission critical, large scale Linux and SRE operations.
  • High visibility role with exposure to senior leadership and engineering stakeholders.
  • Ability to shape operational strategy, automation, and reliability practices.
  • Strong focus on career growth, learning, and leadership development.
  • Comprehensive medical and healthcare plans.
  • In addition to company holidays, ETO and FTO Programs.
  • Maternity and paternity leave, parenting resources, adoption and surrogacy assistance.
  • Purchase Synopsys common stock at a 15% discount, with a 24 month look-back (ESPP).
  • Retirement plans that vary by region and country.
  • Competitive salaries.
#IT#Linux#SRE#operations#automation#reliability#infrastructure#chip design
Synopsys Inc Logo

Company

Synopsys Inc

Job Posted

1 week ago

Employment Type

Full Time

WorkMode

On Site

Experience Level

Senior

Locations

Noida, India

Applicants

Be an early applicant