| Onehouse: "We have observed these improvements leading to 2-10x improvements in query performance for customer tables, … |
@Onehousehq |
Company |
Original |
2026-09-08 |
31 |
2 |
0 |
1 |
0 |
0 |
| CDC-to-lakehouse solutions fall into five categories, and most stop before the last one. https://t.co/sBgN4uJYSD |
@Onehousehq |
Company |
Original |
2026-09-07 |
64 |
3 |
0 |
1 |
0 |
0 |
| Compute-hour billing creates a structural disincentive to ship performance work — for the vendor billing by the hour. h… |
@Onehousehq |
Company |
Original |
2026-09-03 |
68 |
1 |
0 |
1 |
0 |
0 |
| AWS EMR 7.12 posts a 32% price/performance improvement over its prior release. Same benchmark suite, rerun against EMR … |
@Onehousehq |
Company |
Original |
2026-09-02 |
63 |
3 |
0 |
1 |
0 |
0 |
| Table Optimizer's clustering strategies span a spectrum: from simple sorting to advanced multi-dimensional Z-Order/Hilb… |
@Onehousehq |
Company |
Original |
2026-09-01 |
65 |
1 |
0 |
1 |
0 |
0 |
| ClickHouse and StarRocks: best-in-class on Engine Design in a comparison of Spark, ClickHouse, Presto, StarRocks, and T… |
@Onehousehq |
Company |
Original |
2026-08-31 |
101 |
3 |
0 |
1 |
0 |
0 |
| 36% of companies leveraging Spark + k8s are also using one of the managed Spark platforms. https://t.co/yu7jqpXIq9 |
@Onehousehq |
Company |
Original |
2026-08-27 |
56 |
1 |
0 |
1 |
0 |
0 |
| S3 Tables offers fully managed Iceberg compaction. In one hands-on test, a writer at a nominal 1GB/min across 100 parti… |
@Onehousehq |
Company |
Original |
2026-08-26 |
56 |
2 |
0 |
1 |
0 |
0 |
| Your Spark dashboard says the job finished. It won't tell you why one stage burned 80% of the runtime, or whether that … |
@Onehousehq |
Company |
Original |
2026-08-25 |
73 |
2 |
0 |
1 |
0 |
0 |
| Open table formats were meant to end lock-in. They helped—but they didn’t finish the job. Lock-in just moved up a layer… |
@Onehousehq |
Company |
Original |
2026-08-24 |
76 |
3 |
0 |
2 |
0 |
0 |
| Quanton optimizes every stage of the pipeline separately: the read, the compute, and the write. Each stage's gain trace… |
@Onehousehq |
Company |
Original |
2026-08-20 |
47 |
1 |
0 |
1 |
0 |
0 |
| Three Cost Analyzer for Apache Spark™ reports, three different workloads, three different numbers. https://t.co/lT01MBA… |
@Onehousehq |
Company |
Original |
2026-08-19 |
65 |
2 |
1 |
1 |
0 |
0 |
| The math behind Iceberg's deleteOrphanFiles at production scale: a Spark job writing 1 GB of data every 5 minutes, assu… |
@Onehousehq |
Company |
Original |
2026-08-18 |
62 |
1 |
0 |
1 |
0 |
0 |
| In Apna's architecture, records that fail the schema validation check stream to a dedicated quarantine table for review… |
@Onehousehq |
Company |
Original |
2026-08-17 |
68 |
1 |
0 |
1 |
0 |
0 |
| One customer ingesting fast-changing blockchain data needed a hard 10‑minute freshness SLA—even as their workload swung… |
@Onehousehq |
Company |
Original |
2026-08-13 |
64 |
1 |
0 |
1 |
0 |
0 |
| Table Optimizer is executed inside a Spark application, deployed through Kubernetes on the customer's AWS EKS or GCP GK… |
@Onehousehq |
Company |
Original |
2026-08-12 |
67 |
4 |
0 |
1 |
0 |
0 |
| SiliconAngle reported nearly 40% of Snowflake customers are also running Databricks, and nearly 50% of Databricks custo… |
@Onehousehq |
Company |
Original |
2026-08-06 |
149 |
4 |
1 |
2 |
0 |
0 |
| Snowflake's OpenFlow, Confluent's TableFlow, and Databricks's LakeFlow each scope ingestion to a single ecosystem. http… |
@Onehousehq |
Company |
Original |
2026-08-04 |
69 |
2 |
0 |
1 |
0 |
0 |
| OneSync's permission translation is bidirectional across four catalogs, not a one-way push from a single canonical sour… |
@Onehousehq |
Company |
Original |
2026-07-29 |
58 |
2 |
0 |
1 |
0 |
0 |
| Postgres data enters the lakehouse through CDC and is served back out through a Postgres-compatible endpoint — open lak… |
@Onehousehq |
Company |
Original |
2026-07-28 |
69 |
3 |
1 |
1 |
0 |
0 |
| Onehouse Notebooks is organized around one workflow: notebooks are for exploration and development; jobs are for produc… |
@Onehousehq |
Company |
Original |
2026-07-24 |
61 |
2 |
0 |
1 |
0 |
0 |
| Quanton-accelerated Iceberg on Spark, measured per benchmark against OSS Spark:
TPC-DS: up to ~3.1x
TPCx-BB: 2x
TPC-DI… |
@Onehousehq |
Company |
Original |
2026-07-22 |
66 |
3 |
0 |
1 |
0 |
0 |
| Conductor's Principal Engineer, Emil Emilov: "With Onehouse, there's a lot of things we don't have to figure out anymor… |
@Onehousehq |
Company |
Original |
2026-07-21 |
65 |
5 |
1 |
1 |
0 |
1 |
| If you need money for Claude tokens, now you know where to look.
A $1M platform spend on Snowflake, Databricks or EMR m… |
@Onehousehq |
Company |
Original |
2026-07-16 |
102 |
5 |
1 |
1 |
0 |
0 |
| Databricks 47.1%. Kubernetes 33.5%. Every cloud-managed Spark platform combined (EMR, Fabric, DataProc) 19.4%. Kubernet… |
@Onehousehq |
Company |
Original |
2026-07-15 |
546 |
6 |
0 |
1 |
0 |
1 |
| What does compaction, cleaning, and clustering look like when you operate at Uber scale?
At OpenXData, Uber engineers … |
@Onehousehq |
Company |
Original |
2026-04-21 |
97 |
2 |
0 |
0 |
0 |
0 |
| @kevinjqliu is on the Apache Iceberg™ PMC and leads Iceberg work for Microsoft OneLake. At #OpenXData, he’ll lay out a … |
@Onehousehq |
Company |
Original |
2026-04-13 |
112 |
2 |
1 |
0 |
0 |
1 |
| @J_ co-created Apache Parquet, Apache Arrow, and OpenLineage. Three projects. Three industry standards.
Parquet at Twi… |
@Onehousehq |
Company |
Original |
2026-04-03 |
367 |
6 |
2 |
0 |
0 |
0 |
| Announcing Quanton Kubernetes Operator for Apache Spark 🚀
33% organizations adopt Spark on Kubernetes, and now you can… |
@Onehousehq |
Company |
Original |
2026-03-24 |
514 |
12 |
7 |
0 |
1 |
1 |
| Onehouse LakeBase™ - The first lakehouse serving layer with database capabilities like indexing and caching. Built for … |
@Onehousehq |
Company |
Original |
2026-02-17 |
236 |
5 |
2 |
1 |
0 |
0 |
| If you've got half-committed writes staring back at you like a crime scene in your data lake we need to talk...
"The A… |
@Onehousehq |
Company |
Original |
2026-01-26 |
175 |
2 |
1 |
0 |
0 |
0 |
| Are you are itching for a fresh start to your Spark ⚡ data platform in 2026? Migrating from a AWS EMR -> Onehouse is… |
@Onehousehq |
Company |
Original |
2026-01-05 |
98 |
2 |
1 |
0 |
0 |
0 |
| How to design your data lakehouse tables for fast queries
Query performance isn't just about your engine – it starts w… |
@Onehousehq |
Company |
Original |
2025-12-10 |
157 |
4 |
1 |
0 |
0 |
1 |
| Cloud object storage performance is deeply influenced by HTTP behavior. In our latest analysis, we show how S3’s depend… |
@Onehousehq |
Company |
Original |
2025-12-04 |
737 |
6 |
1 |
1 |
1 |
0 |
| With the newly minted AWS EMR 7.12 now generally available, we reran all of the benchmarks from our previous blog, to o… |
@Onehousehq |
Company |
Original |
2025-12-02 |
264 |
5 |
3 |
0 |
0 |
0 |
| ⏰ We’re going live in just a few minutes!
The Open Source Data Summit 2025 kicks off shortly. There's still time to gr… |
@Onehousehq |
Company |
Original |
2025-11-13 |
195 |
1 |
0 |
0 |
0 |
0 |
| Quanton accelerated Iceberg workloads now at 3x performance vs OSS Spark and up to 5x price/performance vs other premiu… |
@Onehousehq |
Company |
Quote |
2025-11-12 |
218 |
3 |
0 |
0 |
0 |
0 |
| Honored to share that Onehouse has been named to @CRN's 2025 Stellar Startups list in Big Data 🎉
We’re excited to see … |
@Onehousehq |
Company |
Original |
2025-11-10 |
115 |
1 |
0 |
0 |
0 |
0 |
| 💸 Most teams running Apache Spark™ are burning 30-70% of their compute budget, and they don’t even know it.
Why? Becau… |
@Onehousehq |
Company |
Original |
2025-11-05 |
192 |
4 |
3 |
0 |
0 |
0 |
| 📘 The full “@apachehudi: The Definitive Guide” is out (free).
Why it matters: Lakehouses are mainstream, but reliable,… |
@Onehousehq |
Company |
Original |
2025-11-04 |
226 |
2 |
2 |
0 |
0 |
0 |
| We’re excited to join the dbt community at #dbtCoalesce next week!
Come by booth #426 to see how Onehouse + dbt make E… |
@Onehousehq |
Company |
Original |
2025-10-10 |
187 |
2 |
0 |
0 |
0 |
0 |
| Are table format wars ⚔️ over yet? Yes, but not in the way you expect. Delta Lake is in the rearview mirror. Apache Hud… |
@Onehousehq |
Company |
Original |
2025-10-02 |
320 |
5 |
3 |
1 |
1 |
2 |
| An uncomfy truth... ⚠️ Most Spark jobs waste compute + eng time
We consistently see 30-70% waste 🧟. Not from lazy engi… |
@Onehousehq |
Company |
Original |
2025-09-25 |
136 |
2 |
2 |
0 |
0 |
0 |
| New work on Generic Table APIs from Snowflake and Onehouse in Polaris uses XTable for non-iceberg tables. Running XTabl… |
@Onehousehq |
Company |
Original |
2025-09-23 |
332 |
4 |
3 |
1 |
0 |
1 |
| 🚨 Spark’s dynamic allocation is broken for modern lakehouse workloads.
It scales by task backlog, not by data volume o… |
@Onehousehq |
Company |
Original |
2025-09-17 |
202 |
4 |
2 |
1 |
0 |
0 |
| 💸 Running @apachehudi on @ApacheSpark or #EMR and watching your costs balloon? You’re not alone.
Small files, sprawlin… |
@Onehousehq |
Company |
Original |
2025-09-10 |
132 |
2 |
0 |
0 |
0 |
0 |
| 🚨 Happening tomorrow! Cut Spark SQL costs 50%+ with dbt + Onehouse.
Don't miss it 👇 |
@Onehousehq |
Company |
Quote |
2025-09-02 |
251 |
2 |
0 |
0 |
0 |
0 |
| Hudi Streamer is your all-in-one tool for building up a data lakehouse. Out of the box, it provides a wide range of dat… |
@Onehousehq |
Company |
Original |
2025-09-02 |
254 |
8 |
0 |
0 |
0 |
1 |
| 💡 Spark pipelines too slow or too expensive?
We just released the Spark Analyzer — a free tool that scans your Spark H… |
@Onehousehq |
Company |
Original |
2025-08-28 |
517 |
4 |
1 |
0 |
0 |
3 |
| 💸 Reality check: ETL writing costs (the "L") can eat 20-50% of your pipeline time, but most benchmarks ignore this comp… |
@Onehousehq |
Company |
Original |
2025-08-11 |
343 |
2 |
1 |
0 |
0 |
0 |
| Plenty of companies are finding out vibe coding has its potential - and its limits. Turns out if you unleash solid engi… |
@Onehousehq |
Company |
Original |
2025-08-07 |
118 |
3 |
0 |
0 |
0 |
0 |
| Everyone is solving ingestion:
❄️ Snowflake – OpenFlow
🧱 Databricks – LakeFlow
📡 Confluent – TableFlow
But they all sha… |
@Onehousehq |
Company |
Original |
2025-08-01 |
620 |
3 |
1 |
2 |
0 |
5 |
| 🎉 A new chapter "Running Hudi in Production" is now available in the early release of "Apache Hudi™: The Definitive Gui… |
@Onehousehq |
Company |
Original |
2025-07-25 |
358 |
3 |
2 |
0 |
0 |
0 |
| 🎉 We're happy to share that 7 chapters are now available in the early release of "Apache Hudi™: The Definitive Guide" -… |
@Onehousehq |
Company |
Original |
2025-07-17 |
557 |
12 |
5 |
0 |
0 |
1 |
| AWS S3 Tables simplifies Iceberg tables, but we benchmarked and discovered:
🔁 3h compaction delays
📉 Perf degradation
… |
@Onehousehq |
Company |
Original |
2025-07-08 |
564 |
7 |
1 |
0 |
1 |
1 |
| If you’re at #DataAISummit or were at #SnowflakeSummit last week, you’ve heard us talking about cutting SQL and Spark … |
@Onehousehq |
Company |
Original |
2025-06-11 |
141 |
2 |
0 |
0 |
0 |
0 |
| At @databricks #DataAISummit? Join us this evening for @Onehousehq VP of Product @KyleJWeller's talk “Open By Default, … |
@Onehousehq |
Company |
Original |
2025-06-10 |
427 |
3 |
1 |
0 |
0 |
0 |
| Who’s going to #databricks #DataAISummit next week? Join us on Tuesday for @Onehousehq VP of Product @KyleJWeller shar… |
@Onehousehq |
Company |
Original |
2025-06-06 |
277 |
4 |
1 |
0 |
0 |
0 |
| It's the last day at #SnowflakeSummit. If you haven't stopped by booth 1415 yet, come down and see what all the noise i… |
@Onehousehq |
Company |
Original |
2025-06-05 |
67 |
2 |
0 |
0 |
0 |
0 |
| Got swag? Swing by booth 1415 and learn what the open #datalakehouse can do to cut your ETL pipeline and data modeling … |
@Onehousehq |
Company |
Original |
2025-06-04 |
88 |
2 |
0 |
0 |
0 |
0 |
| It may be the season for SNOW, but this talk was hot! It was standing room only for @KyleJWeller's presentation on bui… |
@Onehousehq |
Company |
Original |
2025-06-03 |
199 |
5 |
2 |
0 |
0 |
0 |
| Onehouse + Snowflake = unquestionably better together. Standing room only for Onehouse VP of Product @KyleJWeller at #… |
@Onehousehq |
Company |
Original |
2025-06-02 |
399 |
2 |
1 |
0 |
0 |
0 |
| At #SnowflakeSummit? @Onehouse VP of Product @KyleJWeller is live in an hour for “Building the Fastest @ApacheIceberg L… |
@Onehousehq |
Company |
Original |
2025-06-02 |
163 |
5 |
1 |
0 |
0 |
0 |
| Building your #datalakehouse with @ApacheIceberg and @Snowflake?
Join Onehouse VP of Product @KyleJWeller at Snowflake… |
@Onehousehq |
Company |
Original |
2025-05-30 |
150 |
3 |
0 |
0 |
0 |
0 |
| What happens when 𝘮𝘢𝘴𝘴𝘪𝘷𝘦 𝘴𝘵𝘳𝘦𝘢𝘮𝘪𝘯𝘨 𝘸𝘰𝘳𝘬𝘭𝘰𝘢𝘥𝘴 meet the reality of maintaining Iceberg metadata at scale?
We just dropp… |
@Onehousehq |
Company |
Original |
2025-05-29 |
270 |
6 |
3 |
0 |
0 |
1 |
| That's a wrap! What an excellent time at #OpenXData today. 👏 Thanks to @confluentinc, @databricks, and @dbt_labs for co… |
@Onehousehq |
Company |
Original |
2025-05-21 |
115 |
4 |
0 |
0 |
0 |
0 |
| 🚨 It’s almost time — #OpenXData kicks off in just a few minutes!
Doors open at 9:00 AM PT, and the first keynote starts… |
@Onehousehq |
Company |
Original |
2025-05-21 |
186 |
2 |
0 |
0 |
1 |
0 |
| Today we announce SQL and Spark jobs powered by our new Quanton execution engine 🚀
Quanton delivers 2-3x price/perform… |
@Onehousehq |
Company |
Original |
2025-05-20 |
426 |
8 |
5 |
0 |
0 |
0 |
| 📢 Just two days to go! #OpenXData is the premier event on open data architectures for data practitioners this year.
As… |
@Onehousehq |
Company |
Original |
2025-05-19 |
125 |
1 |
2 |
0 |
0 |
0 |
| Cloud data costs are rising, and over 50% of that spend goes into ETL workloads. But how do you 𝘢𝘤𝘤𝘶𝘳𝘢𝘵𝘦𝘭𝘺 measure the … |
@Onehousehq |
Company |
Original |
2025-05-15 |
264 |
2 |
1 |
0 |
0 |
0 |
| 🕒 Ever wanted to spin up a #datalakehouse but couldn't find the time?
⚡ Let Chandra Krishnan, Solutions Engineer at On… |
@Onehousehq |
Company |
Original |
2025-05-12 |
132 |
2 |
1 |
0 |
0 |
0 |
| 🔥 Announcing OpenXData - the free virtual conference on open data 🔥
OpenXData brings together 25+ sessions by data inn… |
@Onehousehq |
Company |
Original |
2025-05-01 |
120 |
3 |
1 |
0 |
0 |
0 |
| We are unveiling #OpenEngines with a live webinar in just over half an hour. @andywalner and @KyleJWeller will share a … |
@Onehousehq |
Company |
Original |
2025-04-29 |
56 |
2 |
0 |
0 |
0 |
0 |
| Cloud warehouses are adopting open table formats like @ApacheIceberg, @apachehudi, and @DeltaLakeOSS — but how open ar… |
@Onehousehq |
Company |
Original |
2025-04-28 |
178 |
2 |
1 |
0 |
0 |
0 |
| 🤔 Scaling ML and DS beyond your laptop? Which path will you choose?
🔹 @ApacheSpark = fast transforms for feature engin… |
@Onehousehq |
Company |
Original |
2025-04-23 |
129 |
4 |
1 |
0 |
0 |
1 |
| Trying to pick the right streaming engine?
Check out our no-fluff breakdown of @ApacheFlink , #KafkaStreams, and #Spar… |
@Onehousehq |
Company |
Original |
2025-04-22 |
149 |
3 |
1 |
0 |
0 |
0 |
| @ClickHouseDB or @StarRocksLabs? @trinodb, @prestodb, or @ApacheSpark?
💡We did the homework so you don’t have to. Full… |
@Onehousehq |
Company |
Original |
2025-04-21 |
184 |
8 |
2 |
0 |
1 |
0 |
| With Open Engines™, you can finally take a ‘horses for courses’ approach—run Flink, Trino, or Ray on your open data, wi… |
@Onehousehq |
Company |
Original |
2025-04-18 |
124 |
3 |
0 |
0 |
0 |
0 |
| 🚨 Announcing Open Engines™, a quick + reliable way to deploy @trinodb, @raydistributed, and @ApacheFlink making it easy… |
@Onehousehq |
Company |
Original |
2025-04-17 |
368 |
8 |
5 |
0 |
2 |
0 |
| We bring the #datalakehouse. You bring your data and compute engine. It's that simple.
👉https://t.co/EqbKqFt4Yv
#data… |
@Onehousehq |
Company |
Original |
2025-04-16 |
124 |
2 |
1 |
0 |
0 |
0 |
| We love all the conversations about open #datalakehouse table formats! But a lakehouse is about more than @apachehudi, … |
@Onehousehq |
Company |
Original |
2025-03-26 |
437 |
5 |
2 |
0 |
0 |
0 |
| At Onehouse, we know the value of real-time insights. That’s why we’re excited to be one of @confluentinc's launch part… |
@Onehousehq |
Company |
Original |
2025-03-21 |
239 |
5 |
1 |
0 |
0 |
0 |
| You’ve settled on @apachehudi , @ApacheIceberg or @DeltaLakeOSS for your open table format, the foundation to your #d… |
@Onehousehq |
Company |
Original |
2025-03-21 |
332 |
5 |
3 |
0 |
0 |
1 |
| New Blog: Tackling Data Duplication in Lakehouse Architectures 🚀
Data duplication could be a silent killer in data pip… |
@Onehousehq |
Company |
Original |
2025-03-20 |
201 |
7 |
2 |
0 |
0 |
1 |
| Lately, we have heard a lot of people conflating open table formats - @apachehudi, @ApacheIceberg, @DeltaLakeOSS - with… |
@Onehousehq |
Company |
Original |
2025-03-13 |
576 |
3 |
3 |
0 |
0 |
1 |
| Format wars are over. The real innovation? Building data platforms that empower teams to use the right tool for their n… |
@Onehousehq |
Company |
Original |
2025-03-12 |
169 |
4 |
1 |
0 |
0 |
0 |
| Data pipelines built on @ApacheHudi and @ApacheSpark are no joke. But they can be quite performant and efficient. Tomor… |
@Onehousehq |
Company |
Original |
2025-02-26 |
269 |
3 |
2 |
0 |
0 |
0 |
| ACID (atomicity, consistency, isolation, and durability) transactions are crucial in data systems for maintaining data … |
@Onehousehq |
Company |
Original |
2025-02-20 |
293 |
2 |
1 |
0 |
0 |
0 |
| Coming tomorrow! |
@Onehousehq |
Company |
Quote |
2025-02-19 |
121 |
3 |
0 |
0 |
0 |
0 |
| Mix and match your table formats: write in @apachehudi, read in @DeltaLakeOSS or @ApacheIceberg .
We've got you cover… |
@Onehousehq |
Company |
Original |
2025-02-13 |
389 |
5 |
2 |
0 |
0 |
1 |
| Getting started with @apachehudi? A few more chapters of @OReillyMedia's definitive guide to Hudi just dropped. Pick up… |
@Onehousehq |
Company |
Original |
2025-02-10 |
461 |
11 |
3 |
1 |
0 |
1 |
| Lambda architectures were never meant to be a final solution, just a temporary bandage until better tech came along to … |
@Onehousehq |
Company |
Original |
2025-01-30 |
93 |
2 |
0 |
0 |
0 |
0 |
| 🏂Snowflake decoupled compute from storage, but did they really decouple compute from storage? 🤔
@andywalner and Ryan G… |
@Onehousehq |
Company |
Original |
2025-01-29 |
118 |
2 |
2 |
0 |
0 |
0 |
| In case you missed it, the lively discussion and Q&A on Onehouse Compute Runtime with @byte_array and @KyleJWeller … |
@Onehousehq |
Company |
Original |
2025-01-24 |
91 |
1 |
0 |
0 |
0 |
0 |
| Clustering is a powerful storage optimization technique in a #lakehouse that directly enhances query performance and re… |
@Onehousehq |
Company |
Original |
2025-01-23 |
106 |
2 |
0 |
0 |
0 |
1 |
| Earlier today, we launched the Onehouse Compute Runtime, a radical rethinking of core lakehouse operations. OCRs featur… |
@Onehousehq |
Company |
Original |
2025-01-16 |
688 |
6 |
3 |
0 |
1 |
0 |
| Today, we launched the Onehouse Compute Runtime (OCR) 🥁, a high-perf 100% Spark compatible runtime designed for unique … |
@Onehousehq |
Company |
Original |
2025-01-16 |
719 |
14 |
5 |
0 |
0 |
0 |