Senior Technical Lead
Teknik, data och digitalt · IT-infrastruktur och säkerhet · DevOps · Site reliability engineering · Molnteknik
I korthet
We are looking for a Senior Technical Lead with over 8 years of experience in DevOps/SRE to design, implement, and manage highly available, scalable, and secure cloud-native platforms. The role involves a reliability-first mindset to drive operational excellence and platform stability, with responsibilities spanning SRE, DevOps automation, Kubernetes management, and cloud operations.
Ansvarsområden
- Design, implement, and maintain reliable, scalable, and resilient infrastructure.
- Define and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and Error Budgets.
- Ensure high availability and performance of mission-critical applications.
- Lead incident management, root cause analysis (RCA), and postmortem reviews.
- Drive reliability improvements through automation and proactive monitoring.
- Build and optimize CI/CD pipelines using modern DevOps tools.
- Automate infrastructure provisioning, deployment, and operational processes.
- Implement Infrastructure as Code (IaC) practices using Terraform, CloudFormation, or similar tools.
- Enable self-service platform capabilities and reduce operational toil through automation.
- Design, deploy, and manage Kubernetes clusters in cloud or hybrid environments.
- Manage containerized workloads using Docker and Kubernetes.
- Implement deployment strategies such as Blue-Green, Canary, and Rolling Deployments.
- Ensure platform security, scalability, and fault tolerance.
- Manage and optimize cloud infrastructure on AWS, Azure, or GCP.
- Monitor cloud resource utilization, availability, and performance.
- Implement cloud security best practices, IAM, RBAC, and governance controls.
- Support disaster recovery and business continuity initiatives.
- Implement monitoring, logging, and alerting solutions.
- Develop dashboards and operational metrics to track platform health.
- Improve MTTR through observability and automation.
- Perform capacity planning and performance tuning.
- Partner with Development, QA, Security, and Architecture teams.
- Promote DevOps, DevSecOps, and SRE best practices across teams.
- Mentor junior engineers and contribute to technical leadership initiatives.
Krav
- 8+ years of experience in DevOps, Platform Engineering, Cloud Operations, or Site Reliability Engineering.
- Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field.
- Strong troubleshooting, debugging, and performance optimization skills.
- Experience supporting large-scale production environments.
- Strong communication, stakeholder management, and cross-functional collaboration skills.
- AWS, Azure, or Google Cloud Platform (GCP)
- Kubernetes
- Docker
- Helm
- Jenkins
- Azure DevOps
- GitHub Actions
- GitLab CI/CD
- Terraform
- Ansible
- CloudFormation (Preferred)
- Prometheus
- Grafana
- Splunk
- ELK / EFK Stack
- AppDynamics or Dynatrace
- Python
- Bash/Shell Scripting
- PowerShell (Preferred)
- Git
- GitHub / GitLab / Bitbucket
- Linux Administration (RHEL, Ubuntu, CentOS)
Önskade kvalifikationer
- Certified Kubernetes Administrator (CKA)
- AWS Certified DevOps Engineer – Professional
- Microsoft Azure DevOps Engineer Expert
- HashiCorp Terraform Associate
- Google Professional Cloud DevOps Engineer
- Certified Kubernetes Security Specialist (CKS)
- OpenShift Administration
- Service Mesh (Istio/Linkerd)
- Kafka/Event Streaming Platforms
- Security & Compliance Automation
- Chaos Engineering
- FinOps and Cloud Cost Optimization
- MLOps Exposure
- DevSecOps Implementation Experience
Förmåner
- Platform Availability of 99.9% or higher
- Reduced Incident Volume and Mean Time to Recovery (MTTR)
- Increased Deployment Frequency and Release Reliability
- Improved Automation Coverage
- Enhanced Infrastructure Security, Scalability, and Reliability
#DevOps#Site Reliability Engineering#SRE#Cloud#Kubernetes#CI/CD#Infrastructure as Code#Automation#AWS#Azure#GCP