cubefs/cubefs

cloud-native distributed storage

What it solves

CubeFS provides a scalable, cloud-native distributed storage solution that eliminates the tight coupling between compute and storage. This allows databases, search systems, and AI/ML applications to scale their processing power and storage capacity independently.

How it works

It functions as a distributed file and object storage system that supports multiple access protocols (POSIX, HDFS, S3, and a native REST API). It utilizes a highly scalable metadata service with strong consistency and offers flexible storage policies, including high-performance replication and low-cost erasure coding. For hybrid cloud setups, it provides I/O acceleration via multi-level caching on top of public cloud storage like S3.

Who it’s for

It is designed for developers and infrastructure engineers building data lakes, datacenter filesystems, or private/hybrid cloud storage, specifically those running AI/ML workloads or large-scale container platforms.

Highlights

  • Supports POSIX, HDFS, and S3 protocols for flexible data access.
  • Enables separation of storage and compute for AI/ML and database architectures.
  • Features multi-tenancy for better resource isolation and utilization.
  • Optimized for both large and small files, as well as sequential and random writes.
  • Provides hybrid cloud acceleration through multi-level caching.

Related

  • Project
  • Project
  • Project
  • Project
  • Project