We’re currently looking for a Senior Ceph/Rook-Ceph + Kubernetes Storage Engineer to join a long-term project. The role focuses on production Ceph environments, Kubernetes storage, and infrastructure troubleshooting.Requirements5+ years in DevOps, SRE, infrastructure or storage engineeringDeep hands-on production Ceph experience: OSD, MON, MGR, PGs, recovery, backfill, capacity planning, performance tuning, scaling and high availabilityExperience building, operating, upgrading and troubleshooting production Ceph clustersHands-on Rook-Ceph in Kubernetes, preferably business-critical production environmentsStrong Kubernetes knowledge, including CSI/storage integrationLinux administration and troubleshooting at OS/hardware levelAutomation experience with AnsibleKubernetes tooling such as HelmMonitoring/observability with Prometheus and GrafanaExperience investigating storage performance, latency, disk failures, network bottlenecks and recovery issuesStrong understanding of failure domains, CRUSH topology, replication and storage architectureGood English communication skillsHighly valuable but not mandatory:OpenStack experience, especially Ceph integration with Cinder, Glance, Nova and RBDPetabyte-scale Ceph environmentsBare-metal infrastructureLarge production Rook-Ceph clustersCephFS and RGW/S3 experienceCustomer-facing troubleshooting/support experience
We’re currently looking for a Senior Ceph/Rook-Ceph + Kubernetes Storage Engineer to join a long-term project. The role focuses on production Ceph environments, Kubernetes storage, and infrastructure troubleshooting.
Requirements
- 5+ years in DevOps, SRE, infrastructure or storage engineering
- Deep hands-on production Ceph experience: OSD, MON, MGR, PGs, recovery, backfill, capacity planning, performance tuning, scaling and high availability
- Experience building, operating, upgrading and troubleshooting production Ceph clusters
- Hands-on Rook-Ceph in Kubernetes, preferably business-critical production environments
- Strong Kubernetes knowledge, including CSI/storage integration
- Linux administration and troubleshooting at OS/hardware level
- Automation experience with Ansible
- Kubernetes tooling such as Helm
- Monitoring/observability with Prometheus and Grafana
- Experience investigating storage performance, latency, disk failures, network bottlenecks and recovery issues
- Strong understanding of failure domains, CRUSH topology, replication and storage architecture
- Good English communication skills
Highly valuable but not mandatory:
- OpenStack experience, especially Ceph integration with Cinder, Glance, Nova and RBD
- Petabyte-scale Ceph environments
- Bare-metal infrastructure
- Large production Rook-Ceph clusters
- CephFS and RGW/S3 experience
- Customer-facing troubleshooting/support experience
Globaldev Group is a team of professionals specializing in creating engineering teams for technological businesses for the Western Europe, Israel, USA. Only long-term product teams for EU with investments and a modern stack of technologies.We set up the company in 2010 in Kharkiv. In 2017 we opened the office in Berlin. Currently, there are 220 of us and we are constantly growing.Now in 2022 we have development hubs in Israel, Ukraine, Portugal and Poland.If you’re interested in joining the team and taking part in decision making that will affect the product and business – let us know!
Globaldev Group is a team of professionals specializing in creating engineering teams for technological businesses for the Western Europe, Israel, USA. Only long-term product teams for EU with investments and a modern stack of technologies.
We set up the company in 2010 in Kharkiv. In 2017 we opened the office in Berlin. Currently, there are 220 of us and we are constantly growing.
Now in 2022 we have development hubs in Israel, Ukraine, Portugal and Poland.
If you’re interested in joining the team and taking part in decision making that will affect the product and business – let us know!