Senior Technical Specialist
Technology, Data & Digital · IT Infrastructure & Security · DevOps · Site Reliability Engineering · Data Engineering · Software Engineering
In short
Join HCLTech as a Senior Technical Specialist in Hyderabad, Telangana. This role focuses on managing Kafka, Confluent Platform, and implementing Site Reliability Engineering principles within a cloud-native environment, specifically on GKE. You will be responsible for operational excellence, automation, and ensuring the reliability of critical banking platforms.
Responsibilities
- Manage Kafka and Confluent Platform, including broker operations, topic design, partitioning, replication, consumer groups, ACLs, security, and certificates.
- Ensure Kafka scalability, performance, and disaster recovery through capacity planning, performance tuning, replication strategies, cluster linking, upgrades, and DR testing.
- Design and implement event-driven architectures and integration patterns, focusing on event design, messaging patterns, consumer onboarding, schema discipline, and asynchronous architectures.
- Provide end-to-end integration support, operational ownership, production readiness, and service health management.
- Implement reliability engineering, incident reduction, automation, and service improvement initiatives.
- Manage logging, alerting, dashboards, tracing, incident response, root cause analysis, and problem management.
- Operate and troubleshoot cloud-native workloads within regulated environments, specifically on Kubernetes (GKE).
- Utilize Infrastructure as Code (Terraform) for infrastructure provisioning, configuration management, and platform automation.
- Implement CI/CD engineering and release automation for build pipelines, deployment automation, and release governance.
- Collaborate across SRE, engineering, and supplier teams to improve service ownership and drive operational maturity.
- Manage TM Vault integration and operations, Apigee API Management, Istio Service Mesh, and cloud database operations (AlloyDB/PostgreSQL).
Requirements
- Mandatory: Kafka & Confluent Platform Administration.
- Mandatory: Kafka Scalability, Performance & Disaster Recovery.
- Mandatory: Event-Driven Architecture & Integration Patterns.
- Mandatory: Site Reliability Engineering (SRE) & Operational Excellence.
- Mandatory: Kubernetes (GKE) Operations & Troubleshooting.
- Mandatory: Infrastructure as Code (Terraform) & Platform Automation.
- Mandatory: CI/CD Engineering & Release Automation.
- Mandatory: Resilience Engineering, Capacity Planning & Recovery Management.
- Mandatory: Cross-Functional Engineering & Supplier Collaboration.
Desired Qualifications
- Preferred: Core Banking Platform & Banking Domain Knowledge.
- Preferred: TM Vault Integration & Operations.
- Preferred: Apigee API Management & Enterprise API Integration.
- Preferred: Istio Service Mesh & Service Communication Architecture.
- Preferred: Distributed tracing, telemetry collection, performance monitoring and operational analytics.
- Preferred: AlloyDB/PostgreSQL & Cloud Database Operations.
- Preferred: Site Reliability Engineering (SRE) Metrics, SLIs, SLOs & Error Budgets.
Benefits
- Supercharge your potential at HCLTech.
- Find your career and spark at a company that puts people first.
- Work at a global technology company with over 223,000 employees in 60 countries.
- Consolidated revenues of $14.8 billion as of 12 months ending June 2026.
#Kafka#Confluent#SRE#Kubernetes#GKE#Terraform#CI/CD#Cloud#DevOps#Banking