Object Storage
S3-compatible storage on Ceph
Storage
We design, build and operate block, file and object storage on commodity hardware: highly available, multi-tenant and entirely as code. We have run Ceph in production since 2016.
The challenge
Data keeps growing, retention rules get stricter and traditional storage ties you to one vendor. At the same time, storage is the layer whose failure takes everything else down with it. Mistakes in failure domains, sizing or networking only show when it matters most.
What we do
Failure domains and data placement, replication or erasure coding, pool and placement-group planning, topologies across several availability zones.
RBD for OpenStack, Kubernetes and hypervisors, NVMe/TCP and iSCSI; CephFS with NFS exports; S3 on RGW with storage classes, versioning and lifecycle rules.
Separation through namespaces, restricted credentials, placement targets and quotas per tenant; a clear model for identities, roles and bucket policies; repeatable onboarding.
Object Lock (WORM), versioning and lifecycle rules as technical controls for retention requirements.
Growth models from usable-to-raw ratio, replication overhead and rebuild reserve; node sizing across CPU, memory, NVMe and network; tuning by measurement, not assumption.
Separate networks for client, cluster, management and out-of-band traffic, port and protocol matrices, MTU requirements, load-balancer and VIP design, S3 endpoints with BGP and ECMP.
Cross-zone replication, active/active load distribution, disaster recovery and business continuity concepts, automated restores that are exercised regularly.
Upgrades and rolling maintenance, adding and draining nodes, rebalancing; diagnosing degraded placement groups, slow requests, drive and network faults.
NetApp ONTAP alongside Ceph and as the starting point for a move to software-defined storage, planned vendor-neutral.
Hardware PoC before procurement, test campaigns for stability, scaling, redundancy and performance, architecture documents traceable to numbered requirements.
Approach
Performance, capacity, availability and compliance are stated in measurable terms.
Hardware and design are tested under real load before anything is bought.
Failure domains, network, capacity and growth are defined and documented.
Unattended installation with cephadm, Ansible and GitOps, including air-gapped environments.
Test campaigns against criteria agreed up front: stability, redundancy, performance.
Monitoring, capacity reports, upgrades and fault analysis, with runbooks for your team.
Technology
Services in detail
S3-compatible storage on Ceph
Highly available volumes on Ceph RBD
Shared file systems with CephFS and NFS
Insights
Why cephadm upgrades stall, why OSDs won’t restart after a version jump, and how to make your next upgrade boring.
Modern storage systems increasingly need to present a highly available and horizontally scalable S3 interface. Ceph RGW has emerged as one of the most flexible S3 gateways available, but to truly...
In dieser Reihe haben wir uns bereits intensiv mit den Grundlagen, der Kostenstruktur und den betrieblichen Herausforderungen von Open‑Source‑Software‑Defined‑Storage für kleine und mittlere...
An architecture review, a new environment from the ground up or support in operations: talk directly to the engineers who will deliver it.