JJobsSonar

Sr. Site Reliability Engineer(Storage Platform)_Remote

Jobs via Dice · United States

SRE / ReliabilityRemote

About this role

The Sr. Site Reliability Engineer (Storage Platform) will be responsible for managing enterprise storage and Kubernetes platforms on Linux. This role requires hands-on experience with SDS solutions such as Ceph and Longhorn, as well as storage migrations from legacy systems. The individual will work with block, file, and object storage, including Fibre Channel and IP-based protocols. Experience with NVMe-oF or iSCSI fabrics is essential, along with expert knowledge of Kubernetes and Linux systems (Ubuntu, RHEL/CentOS). The role also involves proficiency with Infrastructure-as-Code tools like Ansible and Terraform, and strong scripting skills in Python and Bash (Golang is a plus). The candidate must have strong working knowledge of Enterprise DNS and integrations with Kubernetes, and experience operating 24x7 mission-critical production environments. Hands-on experience with KVM hypervisors such as Suse Harvester and OpenStack is required. Strong written and verbal communication skills are also necessary, along with proficiency with Git, CI/CD pipelines, and automated testing frameworks.

Skills & technologies

Must have

  • Ceph
  • Longhorn
  • Kubernetes
  • Linux
  • Ubuntu
  • RHEL/CentOS
  • Ansible
  • Terraform
  • Python
  • Bash
  • Git
  • CI/CD
  • KVM
  • Suse Harvester
  • OpenStack

Nice to have

  • OpenStack Cinder
  • Rubrik
  • CIS/NIST
  • ITIL
  • CKA
  • CKS
  • EX125

Read full description

About the job Dice is the leading career destination for tech experts at every stage of their careers. Our client, Prudent Technologies and Consulting, is seeking the following. Apply via Dice today! Sr. Site Reliability Engineer (Storage Platform) _Remote Contract to-Hire Must Have 6+ years of experience managing enterprise storage and Kubernetes platforms on Linux. Strong hands-on experience with SDS solutions (Ceph, Longhorn) and storage migrations from legacy systems. Experience with block, file, and object storage, including Fibre Channel and IP-based protocols. Experience with NVMe-oF or iSCSI fabrics. Expert knowledge of Kubernetes and Linux systems (Ubuntu, RHEL/CentOS). Proficiency with Infrastructure-as-Code (IaC) (Ansible, Terraform). Strong scripting skills in Python and Bash (Golang (GO) a plus). Strong working knowledge of Enterprise DNS and integrations with Kubernetes Experience operating 24x7 mission-critical production environments. Hands-on experience with KVM hypervisors (Suse Harvester, OpenStack). Strong written and verbal communication skills. Proficiency with Git, CI/CD pipelines, and automated testing frameworks Nice to Have OpenStack Cinder multi-backend administration. Backup platforms (Rubrik). Understanding of CIS/NIST security and infrastructure lifecycle management. ITIL Foundation/advanced certifications in support of ITSM standard methodology. CNCF Certified Kubernetes Administrator (CKA), Certified Kubernetes Security Specialist (CKS) or Red Hat specialist in Ceph Storage Administrator (EX125) certifications.
Ready to apply?Apply now

Similar SRE / Reliability jobs

All SRE / Reliability jobs

Site Reliability Engineer Sr

Dayforce · United States

SRE / ReliabilityRemoteEasy apply$80.5K/yr - $143.8K/yr2w ago

Senior SRE (Cloud)

Hazelcast · United Kingdom

SRE / ReliabilityRemoteEasy apply1mo ago