
Senior Engineering Manager, Site Reliability
Ladders · United States
Remote
About the job
For our client, we are seeking a Senior Engineering Manager Site Reliability to join the team of a leader in the Finance & Insurance space. This role will lead technical work focused on cloud-enabled scalability, reliability, and delivery excellence. You will work across engineering, product, operations, and business stakeholders to translate complex requirements into practical technology solutions. The position offers the opportunity to influence architecture, execution quality, and the technology capabilities that enable long-term growth within a regulated financial services environment.
Location: Remote - US based candidates only, no visa sponsorship available
Compensation: $195,300 – $270,400 annually
Responsibilities
Lead a team in incident management, observability, operational readiness, and reliability engineering
Define and prioritize the SRE function's roadmap and measurable outcomes
Translate reliability strategy into actionable plans with clear ownership
Maintain oversight of team performance and intervene in execution when necessary
Establish a resilient operating model with effective delegation and ownership
Set high technical quality and communication standards
Develop engineers to independently manage complex reliability initiatives
Qualifications
5+ years in reliability engineering management and 7+ years in software, SRE, infrastructure, or platform engineering
Hands-on experience in SRE or similar roles focusing on production systems
Direct management experience in reliability functions with ownership of strategy and outcomes
Strong technical expertise in distributed systems and cloud infrastructure
Experience in leading high severity incident responses at scale
Ability to translate strategic goals into measurable plans
Proven track record in hiring and developing high-performance engineering teams
Benefits
Competitive compensation including bonuses and equity grants
Robust retirement benefits with an attractive company match
Comprehensive health coverage including medical, dental, and vision
Generous paid family leave and support for caregiving needs
Employee Assistance Program focusing on mental health resources
Annual allowances for wellness and productivity tools
Connection and community through team events and employee resource groups
Our client is an equal opportunity employer. We encourage you to apply even if you don’t meet every qualification—your background could be exactly what this team needs.
Desired Skills and Experience
Site Reliability Engineering (SRE), Production Engineering, Incident Management, Observability, Cloud Infrastructure, Distributed Systems, Datadog, Grafana, Prometheus, OpenTelemetry, Kubernetes, AWS, Service-Level Objectives, Error-Budget Practices, Monitoring Tools, Alerting Systems, Incident Response Frameworks, Automation Platforms, Resilience Testing Tools, Configuration Management Tools, Development Operations (DevOps) Practices
$195.3K/yr - $270.4K/yrApply now