
Senior Site Reliability Engineer
Ladders · United States
Remote
About the job
For our client, we are seeking a Senior Site Reliability Engineer to join the team of a leader in the Enterprise Technology space. This role will lead technical work focused on cloud-enabled scalability, reliability, and delivery excellence. You will work across engineering, product, operations, and business stakeholders to translate complex requirements into practical technology solutions. The position offers the opportunity to influence architecture, execution quality, and the technology capabilities that enable long-term growth within a technology-driven environment.
Location: Remote - US based candidates only, no visa sponsorship available
Compensation: $140,000 – $182,000 annually
Responsibilities
Enhance infrastructure automation to reduce manual tasks and mitigate incidents
Develop and maintain core applications for service reliability
Monitor system performance to identify and troubleshoot operational issues
Ensure continuous service delivery through team collaboration
Proactively identify and resolve systemic issues and opportunities for improvement
Qualifications
5-7 years in Site Reliability or Software Engineering with resilient services
Expertise in creating automation and tools for performance management
4+ years of experience with cloud platforms like AWS/GCP
Hands-on experience with observability tools, particularly Prometheus and Grafana
Proficient in on-call duty management and incident response
Strong skills in infrastructure as code, especially with Terraform
Kubernetes operational knowledge, including management of clusters and best practices
Benefits
Comprehensive medical coverage
Financial bonuses potential
Supportive team environment for professional growth
Access to the latest technologies in cloud and SRE
Our client is an equal opportunity employer. We encourage you to apply even if you don’t meet every qualification—your background could be exactly what this team needs.
Desired Skills and Experience
AWS, GCP, Prometheus, Thanos, Grafana, Loki, Tempo, Terraform, Kubernetes, EKS, GKE, NodeJS, Go, Ruby, Python, Linux, Observability standards, SLO frameworks, Shell Scripting, Infrastructure as Code, Monitoring and Alerting Systems, Run Books, Cloud Native Applications, Container Orchestration, Chaos Engineering
$140K/yr - $182K/yrApply now