Home / Companies / ClickHouse / Blog / Post Details
Content Deep Dive

How we built our Internal Data Warehouse at ClickHouse: A year later

Blog post from ClickHouse

Post Details
Company
Date Published
Author
Mihir Gokhale
Word Count
1,958
Company Posts That Month
8
Language
English
Hacker News Points
-
Post removed?
No
Summary

The blog post discusses how ClickHouse's internal data warehouse (DWH) has evolved over the past year to support a more diverse set of users, data sources, and access points. It highlights the use of ClickHouse and dbt as primary components in the stack that have enabled DWH to support real-time data processing into regular batch reporting. The post also covers how the architecture of DWH has been configured with nineteen raw data sources, handling 6 billion rows and 50 TBs of data daily. It mentions how dbt centralized transformation logic related to batch reporting in one place, making it easier to manage growing complexity as SQL became the way business logic was encoded. The post also discusses incorporating more real-time data into DWH and configuring additional access points for users. Finally, it outlines future plans for scaling DWH by decentralizing compute resources and exploring AI features.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Real-time 12 3,932 887 192 +47%
Data Pipeline 3 1,400 332 68 +111%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.