Training

Ceph course

Ceph is not learned by reading the documentation. It is learned by seeing what happens when a node dies, when space runs out, when recovery never finishes. That is where the course starts.

Who it is for

Administrators with a Ceph cluster to run, people evaluating whether it makes sense for them, and anyone handed a cluster somebody else built. Linux and networking are assumed; Ceph is not.

Two levels

Foundation

How it is built and how to stand it up.

  • architecture: monitors, managers, OSDs and what each one does
  • how Ceph decides where data goes, and why that matters
  • replication and erasure coding: what to choose and what it costs
  • the network, which is what separates a cluster that performs from one that struggles
  • installation, inside Proxmox VE and standalone
  • the three ways to use it: disks for virtual machines, shared files, object storage

Advanced

Keeping it up when something goes wrong.

  • reading cluster health and working out what is actually happening
  • disks dropping in and out: it is almost never the disk
  • data that cannot finish syncing, and the order things get fixed in
  • a cluster near full: the most delicate situation there is
  • recovery after a failure: why it is slow and when it can be pushed
  • replacing disks and nodes, growing without stopping anything
  • performance: where it is lost and how it is measured

The programme adapts. The outline above is a starting point. If you already run a cluster we start from that one: we look at how it is built, what worries you, and the course follows your questions rather than an index.

Why us

We hold the Red Hat certification on Ceph Cloud Storage and work on clusters in emergency, most of them built by somebody else. The situations we bring into the room are those: not constructed examples, but real clusters that had stopped and were brought back.

What we work with

  • Ceph in production
  • Red Hat Certified Specialist — Ceph Cloud Storage

How it runs

In our classroom

A real cluster to break and put back together.

At your site

On your cluster, with your data and your real limits.

Online

When people are in different places, or time is short.

There is a Proxmox VE course too

From the first installation to a highly available cluster, taught by people who run them in production.

Go to the Proxmox VE course

Want to arrange a course?

Tell us how many people, whether you already run a cluster, and what you need to understand. We will propose a programme and a length.

Get in touch