Senior Site Reliability Engineer Lead
Technology, Data & Digital · IT Infrastructure & Security · Site Reliability Engineering · DevOps
In short
HCLTech is seeking a Senior Site Reliability Engineer Lead in Bangalore, India, to manage a team of support engineers, ensure system uptime, and drive automation. This role requires strong experience with GitHub Enterprise, GCP, CI/CD pipelines, and scripting, along with production support and incident management skills.
Responsibilities
- Lead and manage a team of support engineers in resolving incidents, requests, and problems to ensure system uptime and reliability.
- Collaborate with engineering and development teams to implement efficient and scalable solutions that enhance system performance.
- Develop and maintain support documentation, standard operating procedures, and best practices for the support team.
- Identify opportunities for automation and implement tools to streamline support processes.
- Monitor system performance and provide recommendations for improvements to optimize system reliability.
- Participate in on call rotations to address critical incidents and ensure 24/7 system availability.
- Conduct regular performance evaluations, provide feedback, and mentor team members to promote professional growth.
Requirements
- Strong hands-on experience with GitHub Enterprise administration and support.
- Good understanding of Git concepts, branching strategies, pull requests, code reviews, and repository administration.
- Experience supporting enterprise source code management platforms such as GitHub, Bitbucket, GitLab or similar tools.
- Experience supporting and troubleshooting source code migrations; Bitbucket to GitHub migration experience is highly preferred, while similar migration experience is also welcome.
- Understanding of CI/CD pipelines and modern developer tooling ecosystems.
- Experience working in Google Cloud Platform (GCP) environments.
- Knowledge of scripting and automation using Python, Shell, PowerShell or similar technologies is beneficial.
- Production support experience for developer platforms, SDLC tools, Git repositories or DevOps platforms.
- Experience handling incidents, service requests, root cause analysis and problem management activities.
- Ability to troubleshoot technical issues independently and provide practical solutions.
- Experience coordinating with global application development teams and stakeholders.
- Strong communication, documentation, stakeholder management and organizational skills.
Desired Qualifications
- Relevant certifications in Site Reliability Engineering (SRE) or Cloud Services are a plus.
Benefits
- Supercharge your potential at HCLTech.
- Find your career.
- Find your spark.
- A place that knows that helping its customers stay on top starts by putting its people first.
#Site Reliability Engineering#GitHub#Git#CI/CD#Google Cloud Platform#Python#Shell#PowerShell#DevOps#SRE