
Senior Engineering Manager, Site Reliability
Ladders · United States
Remote
About the job
For our client, we are seeking a Senior Engineering Manager Site Reliability to join the team of a leader in the Telecommunications & Hardware space. This role will lead technical work focused on cloud-enabled scalability, reliability, and delivery excellence. You will work across engineering, product, operations, and business stakeholders to translate complex requirements into practical technology solutions. The position offers the opportunity to influence architecture, execution quality, and the technology capabilities that enable long-term growth within a Telecommunications & Hardware environment.
Location: Remote - US based candidates only, no visa sponsorship available
Compensation: $260,000 – $280,000 annually
Responsibilities
Establish and build a site reliability engineering team from scratch
Define and professionalize incident management processes across engineering teams
Promote incident professionalism and a reliability culture through training
Drive engineering excellence with design and code reviews and retrospectives
Manage a balance between incident response and a reliability engineering roadmap
Recruit, mentor, and develop site reliability engineers
Collaborate with other leaders to foster ownership and define SLOs
Qualifications
5-7 years of experience in leading SRE or Infrastructure teams
Hands-on experience as a Site Reliability Engineer with familiarity in operational processes
Expertise in observability tools and metrics such as APM and service-specific metrics
Experience in incident management tooling and vendor decision making
Proficiency with at least one major cloud provider, preferably AWS
Benefits
Inclusive and diverse team culture
Opportunities for career development and advancement
Collaborative and innovative work environment
Supports hybrid and remote work models
Comprehensive health, vision, and dental insurance for employees and families
Our client is an equal opportunity employer. We encourage you to apply even if you don’t meet every qualification—your background could be exactly what this team needs.
Desired Skills and Experience
Site Reliability Engineering (SRE), Incident Management Tools (PagerDuty, FireHydrant), Cloud Infrastructure (AWS, GCP, Azure), Application Performance Management (APM) tools, Runbook Development Tools, Metrics and Monitoring (Prometheus, Grafana), Infrastructure as Code (IaC) tools (e.g., Terraform, CloudFormation) - likely needed for deploying and managing infrastructure efficiently in cloud environments; DevOps tools and practices - considering the emphasis on engineering culture and excellence; Agile project management tools (e.g., Jira, Trello) - important for managing the development and operational backlog in a team setup.
$260K/yr - $280K/yr · Dental benefitApply now