Senior Technical Lead

Teknik, data och digitalt · IT-infrastruktur och säkerhet · DevOps · Site reliability engineering · Molnteknik

I korthet

We are looking for a Senior Technical Lead with over 8 years of experience in DevOps/SRE to design, implement, and manage highly available, scalable, and secure cloud-native platforms. The role involves a reliability-first mindset to drive operational excellence and platform stability, with responsibilities spanning SRE, DevOps automation, Kubernetes management, and cloud operations.

Ansvarsområden

  • Design, implement, and maintain reliable, scalable, and resilient infrastructure.
  • Define and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and Error Budgets.
  • Ensure high availability and performance of mission-critical applications.
  • Lead incident management, root cause analysis (RCA), and postmortem reviews.
  • Drive reliability improvements through automation and proactive monitoring.
  • Build and optimize CI/CD pipelines using modern DevOps tools.
  • Automate infrastructure provisioning, deployment, and operational processes.
  • Implement Infrastructure as Code (IaC) practices using Terraform, CloudFormation, or similar tools.
  • Enable self-service platform capabilities and reduce operational toil through automation.
  • Design, deploy, and manage Kubernetes clusters in cloud or hybrid environments.
  • Manage containerized workloads using Docker and Kubernetes.
  • Implement deployment strategies such as Blue-Green, Canary, and Rolling Deployments.
  • Ensure platform security, scalability, and fault tolerance.
  • Manage and optimize cloud infrastructure on AWS, Azure, or GCP.
  • Monitor cloud resource utilization, availability, and performance.
  • Implement cloud security best practices, IAM, RBAC, and governance controls.
  • Support disaster recovery and business continuity initiatives.
  • Implement monitoring, logging, and alerting solutions.
  • Develop dashboards and operational metrics to track platform health.
  • Improve MTTR through observability and automation.
  • Perform capacity planning and performance tuning.
  • Partner with Development, QA, Security, and Architecture teams.
  • Promote DevOps, DevSecOps, and SRE best practices across teams.
  • Mentor junior engineers and contribute to technical leadership initiatives.

Krav

  • 8+ years of experience in DevOps, Platform Engineering, Cloud Operations, or Site Reliability Engineering.
  • Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field.
  • Strong troubleshooting, debugging, and performance optimization skills.
  • Experience supporting large-scale production environments.
  • Strong communication, stakeholder management, and cross-functional collaboration skills.
  • AWS, Azure, or Google Cloud Platform (GCP)
  • Kubernetes
  • Docker
  • Helm
  • Jenkins
  • Azure DevOps
  • GitHub Actions
  • GitLab CI/CD
  • Terraform
  • Ansible
  • CloudFormation (Preferred)
  • Prometheus
  • Grafana
  • Splunk
  • ELK / EFK Stack
  • AppDynamics or Dynatrace
  • Python
  • Bash/Shell Scripting
  • PowerShell (Preferred)
  • Git
  • GitHub / GitLab / Bitbucket
  • Linux Administration (RHEL, Ubuntu, CentOS)

Önskade kvalifikationer

  • Certified Kubernetes Administrator (CKA)
  • AWS Certified DevOps Engineer – Professional
  • Microsoft Azure DevOps Engineer Expert
  • HashiCorp Terraform Associate
  • Google Professional Cloud DevOps Engineer
  • Certified Kubernetes Security Specialist (CKS)
  • OpenShift Administration
  • Service Mesh (Istio/Linkerd)
  • Kafka/Event Streaming Platforms
  • Security & Compliance Automation
  • Chaos Engineering
  • FinOps and Cloud Cost Optimization
  • MLOps Exposure
  • DevSecOps Implementation Experience

Förmåner

  • Platform Availability of 99.9% or higher
  • Reduced Incident Volume and Mean Time to Recovery (MTTR)
  • Increased Deployment Frequency and Release Reliability
  • Improved Automation Coverage
  • Enhanced Infrastructure Security, Scalability, and Reliability
#DevOps#Site Reliability Engineering#SRE#Cloud#Kubernetes#CI/CD#Infrastructure as Code#Automation#AWS#Azure#GCP
HCLTech Logo

Företag

HCLTech

Publicerade jobb

för 1 månad sedan

Anställningstyp

Heltid

Arbetsform

På plats

Erfarenhetsnivå

Senior

Platser

Bangalore, India

Kvalifikation

Kandidatexamen

Sökande

Ansök tidigt