HCLTech – Hungary

SRE Technical Specialist

HCLTech – Hungary  •  Onsite  •  2 hours ago
Apply
AI can make mistakes so check important info. Chat history is never stored.

Job Description

Job Summary

Seeking an experienced Support Specialist having experience of 10+ years working as Site Reliability Engineer (SRE) with strong expertise in Kubernetes, Linux, Cloud Platforms (AWS/Azure/GCP), Observability, Automation, and Production Support. Responsible for managing and supporting production Kubernetes environments, ensuring platform reliability, availability, security, scalability, and operational excellence.

Key Responsibilities:

  • Manage and maintain Kubernetes clusters, including deployments, upgrades, capacity planning, RBAC, networking, storage, Ingress, ConfigMaps, Secrets, Services, Persistent Volumes, StatefulSets, DaemonSets, Jobs/CronJobs, Helm, and Autoscaling (HPA/VPA).
  • Provide 24x7 production support, incident management, RCA, postmortems, service restoration, and SLA/SLO compliance.
  • Monitor and improve platform reliability using Prometheus, Grafana, Loki, Elastic Stack, OpenTelemetry, and AlertManager.
  • Troubleshoot Kubernetes, Linux, container (Docker/OCI), networking (DNS, Load Balancers, TLS, Ingress), cloud, and infrastructure issues.
  • Automate operational tasks through Bash, Python, Terraform, and Ansible, and support Infrastructure as Code practices.
  • Support CI/CD and release management using GitHub Actions, GitLab CI, Jenkins, and ArgoCD (preferred).
  • Perform patching, cluster maintenance, security updates, backups, disaster recovery validation, and platform upgrades.
  • Create runbooks, operational documentation, dashboards, alerts, and capacity planning reports.
  • Collaborate with Development, Platform Engineering, Security, Networking, Cloud Operations, and DevOps teams to improve system resilience and operational efficiency.

Required Skills: Kubernetes Administration, Linux, Docker/OCI, AWS/Azure/GCP, Networking, CI/CD, GitOps, Observability, Incident Management, RCA, Automation, Terraform, Ansible, Bash, Python.

Preferred: CKA/CKS certification, Cloud certifications, Multi-cluster/Multi-region Kubernetes, Service Mesh (Istio/Linkerd), High Availability, Disaster Recovery, Security Hardening, Capacity Planning, Performance Tuning, Cost Optimization, Chaos Engineering, AI-assisted Observability.

Key Competencies: Strong troubleshooting, ownership, production support, customer focus, communication, collaboration, continuous improvement, and ability to perform under pressure.

Success Metrics: High platform availability, improved MTTR, reduced incidents and alert noise, SLA/SLO compliance, increased automation coverage, successful upgrades/maintenance, and customer satisfaction.

Key Responsibilities

null

Skill Requirements

null

Other Requirements

Preferred Qualifications

  • Certified Kubernetes Administrator (CKA)
  • Certified Kubernetes Security Specialist (CKS)
  • Cloud certifications (AWS/Azure/GCP)
  • Experience supporting multi-cluster Kubernetes environments.
  • Experience with service mesh technologies (Istio/Linkerd).
  • Experience with GitOps workflows.
HCLTech – Hungary

About HCLTech – Hungary

HCLTech is a global technology company, home to more than 226,600 people across 60 countries, delivering industry-leading capabilities centered around digital, engineering, cloud and AI, powered by a broad portfolio of technology services and products. We work with clients across all major verticals, providing industry solutions for Financial Services, Manufacturing, Life Sciences and Healthcare, Technology and Services, Telecom and Media, Retail and CPG, and Public Services. Consolidated revenues as of 12 months ending September 2025 totaled $14.2 billion. To learn how we can supercharge progress for you, visit hcltech.com.

Industry
IT & Software
Company Size
51-200 employees
Headquarters
Budapest, HU
Year Founded
2006
Social Media