Senior Ceph/Rook-Ceph + Kubernetes Storage Engineer
- Role
- DevOps
- Experience
- Senior
Open to RO, PL, BG, UA, SK only. Set where you work from to check your eligibility.
No BS summary
Senior storage infrastructure engineer, 5+ yrs in DevOps/SRE/infra. Must have deep hands-on production Ceph (OSD, MON, recovery, performance tuning) plus Rook in Kubernetes, K8s CSI/storage, Linux, Ansible, Prometheus/Grafana. Good English required; being based in the listed central/eastern EU countries.
Core skills
Required skills
Optional skills
Required languages
We’re currently looking for a Senior Ceph/Rook-Ceph + Kubernetes Storage Engineer to join a long-term project. The role focuses on production Ceph environments, Kubernetes storage, and infrastructure troubleshooting. Requirements 5+ years in DevOps, SRE, infrastructure or storage engineering Deep hands-on production Ceph experience: OSD, MON, MGR, PGs, recovery, backfill, capacity planning, performance tuning, scaling and high availability Experience building, operating, upgrading and troubleshooting production Ceph clusters Hands-on Rook-Ceph in Kubernetes, preferably business-critical production environments Strong Kubernetes knowledge, including CSI/storage integration Linux administration and troubleshooting at OS/hardware level Automation experience with Ansible Kubernetes tooling such as Helm Monitoring/observability with Prometheus and Grafana Experience investigating storage performance, latency, disk failures, network bottlenecks and recovery issues Strong understanding of failure domains, CRUSH topology, replication and storage architecture Good English communication skills Highly valuable but not mandatory: OpenStack experience, especially Ceph integration with Cinder, Glance, Nova and RBD Petabyte-scale Ceph environments Bare-metal infrastructure Large production Rook-Ceph clusters CephFS and RGW/S3 experience Customer-facing troubleshooting/support experience
What you'll do
- Focus on production Ceph environments, Kubernetes storage, and infrastructure troubleshooting
What they require
- 5+ years in DevOps, SRE, infrastructure or storage engineering
- Deep hands-on production Ceph experience: OSD, MON, MGR, PGs, recovery, backfill, capacity planning, performance tuning, scaling and high availability
- Experience building, operating, upgrading and troubleshooting production Ceph clusters
- Hands-on Rook-Ceph in Kubernetes, preferably business-critical production environments
- Strong Kubernetes knowledge, including CSI/storage integration
- Linux administration and troubleshooting at OS/hardware level
- Automation experience with Ansible, Kubernetes tooling such as Helm
- Monitoring/observability with Prometheus and Grafana
- Experience investigating storage performance, latency, disk failures, network bottlenecks and recovery issues
- Strong understanding of failure domains, CRUSH topology, replication and storage architecture