
Site Reliability Engineer
Hays · European Union
Remote
About the job
Azure Site Reliability Engineer
Location: Remote (Poland & Central/Eastern Europe preferred)
Language: English
Contract: Full Time
We are supporting a major international organisation looking for several Azure Site Reliability Engineers to join a large-scale cloud operations and reliability programme.
Key Responsibilities
Ensure reliability, availability and performance of cloud-native platforms running on Microsoft Azure
Manage and support Azure Kubernetes Service (AKS)
Implement Infrastructure as Code using Terraform
Build and optimise CI/CD pipelines using GitHub Actions and/or Jenkins
Monitor application and infrastructure health using modern observability tools
Investigate incidents and perform root cause analysis (RCA)
Develop dashboards, alerts and monitoring solutions
Collaborate with engineering and development teams
Create operational runbooks and automation solutions
Contribute to platform stability, scalability and continuous improvement
Required Skills
8+ years of experience in Cloud Infrastructure, DevOps or Site Reliability Engineering
Strong hands-on experience with Microsoft Azure
Proven experience managing Kubernetes environments, including AKS
Experience with Terraform and Infrastructure as Code
Experience with GitHub Actions and/or Jenkins
Scripting skills in Python, Bash, Shell or PowerShell
Experience with monitoring and observability solutions
Strong troubleshooting and incident management skills
Fluent English
Nice to Have
Grafana
Prometheus
Loki
Azure Monitor
Ansible
IoT environments
SLI / SLO / SLA implementation
Azure certifications (AZ-400, Azure Solutions Architect)
What We Offer
Fully remote project
Long-term assignment
International environment
Modern Azure and Kubernetes ecosystem
Opportunity to work on large-scale enterprise platforms
Ready to apply?Apply now