
Sr. Site Reliability Engineer(Storage Platform)_Remote
Jobs via Dice · United States
About this role
The Sr. Site Reliability Engineer (Storage Platform) will be responsible for managing enterprise storage and Kubernetes platforms on Linux. This role requires hands-on experience with SDS solutions such as Ceph and Longhorn, as well as storage migrations from legacy systems. The individual will work with block, file, and object storage, including Fibre Channel and IP-based protocols. Experience with NVMe-oF or iSCSI fabrics is essential, along with expert knowledge of Kubernetes and Linux systems (Ubuntu, RHEL/CentOS). The role also involves proficiency with Infrastructure-as-Code tools like Ansible and Terraform, and strong scripting skills in Python and Bash (Golang is a plus). The candidate must have strong working knowledge of Enterprise DNS and integrations with Kubernetes, and experience operating 24x7 mission-critical production environments. Hands-on experience with KVM hypervisors such as Suse Harvester and OpenStack is required. Strong written and verbal communication skills are also necessary, along with proficiency with Git, CI/CD pipelines, and automated testing frameworks.
Skills & technologies
Must have
- Ceph
- Longhorn
- Kubernetes
- Linux
- Ubuntu
- RHEL/CentOS
- Ansible
- Terraform
- Python
- Bash
- Git
- CI/CD
- KVM
- Suse Harvester
- OpenStack
Nice to have
- OpenStack Cinder
- Rubrik
- CIS/NIST
- ITIL
- CKA
- CKS
- EX125