JJobsSonar

Senior SRE / DevOps Engineer

Confidential · United States

SRE / ReliabilityRemote

About this role

The Senior SRE / DevOps Engineer will work remotely to scale the reliability, performance, and operational maturity of a cloud platform. This hands-on role focuses on improving system reliability, strengthening platform foundations, and enabling engineering teams to move faster safely. The engineer will contribute to infrastructure improvements, technical reviews, observability, and incident support, while also collaborating with product engineering teams.

Skills & technologies

Must have

  • AWS
  • Kubernetes
  • Datadog
  • Terraform
  • CI/CD
  • Cloudflare
  • Observability
  • System Performance Analysis
  • Security Fundamentals
  • Infrastructure Design

Nice to have

  • Go
  • Postgres
  • Cloudflare

Read full description

About the job Senior SRE / DevOps Engineer A US based startup is confidentially hiring a fully remote senior-level SRE / DevOps engineer to help scale the reliability, performance, and operational maturity of a rapidly growing cloud platform. This is a highly hands-on infrastructure role focused on improving system reliability, strengthening platform foundations, and enabling engineering teams to move faster safely. The position is not a traditional people-management or feature-delivery role. This engineer will operate as a senior technical contributor who raises the operational bar across the organization through infrastructure improvements, technical reviews, observability work, and incident support. What You'll Do Reliability and Platform Engineering Improve platform availability and operational resilience Drive improvements around RTO, system performance, and operational recovery Identify and remediate reliability bottlenecks across infrastructure and deployment systems Strengthen observability, alerting quality, and operational tooling Kubernetes & Cloud Infrastructure Maintain and improve production Kubernetes infrastructure Enhance deployment and CI/CD systems, including GitHub Actions and Argo CD workflows Build and evolve reusable Terraform modules and infrastructure patterns Contribute to AWS and Cloudflare infrastructure architecture and reliability Engineering Enablement Review infrastructure and operational changes across teams Unblock engineers during incidents or complex operational work Raise engineering standards around reliability, observability, and operational readiness Partner with product engineering teams without becoming embedded in day-to-day feature pairing On-Call Expectations Engineers in this role may be pulled in to assist during complex incidents Participation in a small platform-focused on-call rotation covering AWS, Cloudflare, Kubernetes, and shared infrastructure issues Required Qualifications 7+ years of experience in SRE, DevOps, Platform Engineering, or related infrastructure roles Demonstrated engineering track record in fully remote, startup environments Comfort using AI / agentic tooling as part of day-to-day engineering workflows, including tools such as Claude Code, Cursor, or similar Currently operating at a Senior Engineer level (or equivalent) 3+ years of deep, hands-on production experience with: AWS Kubernetes Datadog Terraform Strong cloud architecture and infrastructure design skills Solid understanding of security fundamentals and cloud security best practices Deep observability, debugging, and system performance analysis experience Nice to Have Experience with Go Postgres operational experience Cloudflare expertise Fintech or other regulated-environment experience
Ready to apply?Apply now

Similar SRE / Reliability jobs

All SRE / Reliability jobs

Site Reliability Engineer Sr

Dayforce · United States

SRE / ReliabilityRemoteEasy apply$80.5K/yr - $143.8K/yr2w ago

Senior SRE (Cloud)

Hazelcast · United Kingdom

SRE / ReliabilityRemoteEasy apply1mo ago