Ionos2 logo
Posted 2h ago•Revaler Straße 28-31, 10245 Berlin

Staff Storage Reliability Engineer (f/m/d)

staffOn-site (Revaler Straße 28)Salary undisclosed
Required Skills
Next.jsDockerLinux
Job Description

About IONOS

At IONOS, we don't just manage servers – we shape the digital future for more than 6.2 million customers worldwide with our solutions. As Europe's leading hosting provider and a pioneer in independent cloud solutions, we are building next-generation infrastructure – from sovereign cloud architectures and high-performance GPU clusters to integrated AI automation tools.

What drives us is digital sovereignty for Europe and real impact for small and medium-sized businesses (SMBs) as well as large enterprises. We work in agile teams, rely on transparent structures, and believe that excellence and innovation only emerge through true team spirit.

Ready for your next step? Become part of IONOS and grow with us.

Staff Storage Reliability Engineer for our global, public Object Storage platform, built upon Ceph.  You'll be engineering solutions to scale in production, working with our operations team to build, deploy and maintain our platforms.

Currently in double digit PetaBytes, deployed in multiple locations, our Ceph platform is a high growth and a critical part of our internal and public infrastructure.  You will be expected to contribute to the ongoing development, improve and maintain our production environments and ensure that uptime, performance and security are maintained with scaling.

Tasks

Ceph Deployment and Management:

  • Deploying, configuring, and maintaining Ceph clusters, including managing storage pools, placement groups, and other core components.
  • Performance Optimization:
  • Tuning Ceph for optimal performance, addressing bottlenecks, and ensuring efficient resource utilization.

Automation:

  • Developing and implementing automation strategies for Ceph deployments, upgrades, and maintenance tasks.
  • Troubleshooting and Problem Solving:
  • Diagnosing and resolving complex technical issues related to Ceph storage, often involving collaboration with other teams.

Collaboration:

  • Working closely with development teams, system administrators, and other stakeholders to integrate Ceph into various systems and applications.

Staying Updated:

  • Keeping abreast of the latest Ceph developments, new features, and best practices.
  • Participating in the Ceph community and sharing knowledge.  

Qualifications

  • 5 years+ experience as a Senior Linux Engineer or Site Reliability Engineer; Deep and broad understanding of Linux systems and networking.
  • Proficiency in Ceph storage architecture and administration.
  • Knowledge of Cloud Storage technologies (File, Object, Block)
  • Experience with automation tools (eg, Ansible), monitoring and observability
  • Familiarity with cloud platforms and containerization technologies (eg, docker).
  • Excellent problem-solving and troubleshooting skills, strong communication and collaboration abilities.

Benefits

  • Hybrid working model.
  • Flexible working hours through trust-based working hours.
  • At some locations a subsidized canteen and various free drinks.
  • Modern office space with very good transport connections.
  • Various employee discounts for activities and products.
  • Employee events such as summer and winter parties, as well as workshops.
  • Numerous training and development opportunities.
  • Various health offers, such as sports and health courses.

Application Note

We value diversity and welcome all applications – regardless of, for example, gender, nationality, ethnic or social origin, religion, disability, age as well as sexual orientation and identity, physical characteristics, marital status or any other irrelevant factor subject to applicable law.

 

Similar Openings in Other

View all in category➔