Home / Companies / Tinybird / Blog / Post Details
Content Deep Dive

How Tinybird's storage architecture works: S3, local caching, and zero-copy replication

Blog post from Tinybird

Post Details
Company
Date Published
Author
Daniel Pozo
Word Count
1,084
Company Posts That Month
34
Language
English
Hacker News Points
-
Post removed?
No
Summary

Tinybird utilizes a modified version of ClickHouse to optimize data processing by employing a compute-storage separation model, where data is stored in AWS S3 or Google Cloud Storage and cached on local SSDs for enhanced speed and efficiency. The architecture supports zero-copy replication, allowing multiple replicas to reference a single data copy, thus reducing storage costs and improving replication speed. Data ingestion into the ClickHouse cluster is achieved through either a streaming process, managed by the Gatherer to batch events for efficient processing, or a direct batch process, with all writes directed to object storage. The system also employs a packed part format to minimize S3 write operations, significantly cutting infrastructure costs for clients with high data ingestion rates. Tinybird manages all underlying infrastructure elements, including the local cache and replication processes, allowing users to focus on data management without handling the complexities of the system architecture.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Real-time 2 6,457 1,307 242 +28%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.