We're starting a book club where we will be reading Designing Data Intensive Applications - 2nd Edition.
Sign up if you're interested -> https://t.co/hfbHCcNM83
Registration closes on Wednesday the 26th at 9 PM IST. Stay tuned for further updates!
On behalf of 1.4 billion Indians, I extend my warmest congratulations to President Trump and the people of the United States on the historic 250th anniversary of your Independence.
India and the United States share more than a strategic partnership. Our shared belief in democracy, rule of law and the limitless potential of our people make our friendship a force for global good.
May the next 250 years bring even greater prosperity, peace and progress for America and take the India-US partnership to new heights.
@POTUS@realDonaldTrump
We benchmarked @ApacheIceberg vs @databricks@DeltaLakeOSS on a 1TB TPC-H workload!
Iceberg + OLake ran 18.2% faster at 61% lower cost, scaling seamlessly to billions of rows.
With open, flexible support for @ApacheSpark, Trino, @ApacheFlink, a real advantage for modern team
Akshay came across india's largest open source event like this..
I’ve always been impressed by open source since my #GSoC days. Managing pull requests and contributing to large-scale repositories taught me how to build better, collaborate effectively, and follow strong industry practices.
Fast forward to today I’m managing my company’s @_olake booth at India’s largest open-source event, @FOSSUnitedBLR@FOSSUnited
Made some great connections and looking forward to even more meaningful contributions to the open-source ecosystem.
#OpenSource
We’re coming to Mumbai for AWS Community Day @awsugmum on Oct 11!
We’re proud to be the sponsor of the Community Day—great to be part of the global stage.
We’re excited to connect and share why we’re the fastest data replication tool in the world for Apache Iceberg
In our recent session with Arsham (co-founder , greybeam), we dove deep into the evolving @ApacheIceberg catalog ecosystem.
One highlight: Apache Polaris
Here’s what it is — and why it matters 👇
#DataEngineeringStudy
Last week I got one of the biggest names in data engineering to talk on one of the most important topics of 2025 → @ApacheIceberg
From co founders to team managers we had it all .
If you’ve been following @ClickHouseDB , you know it’s already huge in analytics.
Now with read + write support for Iceberg, the ecosystem has evolved further and that's for what we had Shivji & Saurabh the "open source experts" from @nutanix walk us through the updates + live demo.
Got 100+regs on this one and all engineers in the Big Tech interested to dive deeper and network in our call .
We also wrapped our @_olake Community Call
⚡️ Helm deployment (big win for easy setup)
⚡️ @Oracle connector demoed by @SchitizS
I hosted this end to end → and we hold such calls monthly!
Finally, a Catalogs Deep Dive with Arsham who joined us all the way from San Francisco
→ breaking down how 2025 is shaping metadata in Iceberg: Glue, Lakekeeper, @apachepolaris , Nessie, plus @Snowflake updates.
100+regs and this call had a long set of questions coming in (recording in thread if interested )
My main aim at OLake is to create a space where people teach, learn, and most importantly question everything. Many of our calls go a long way in shaping real decisions for teams.
And while it’s not all about tech, the active discussions are integral and range from
"how's your day been ? "
to ... "which catalog are you using in your company and what are the cost savings? "
One thing that stood out recently: Iceberg adoption is rising fast not just in India and China, but all the way in San Francisco too. 🌍
If you’re interested in networking, connecting, and be a part of this ever growing industry, OLake is where we bridge that gap.
Join our Slack to stay updated link in thread 👇
#dataengineering #tech #StartupLife #community #CommunityAmbassador
Medallion Architecture, explained by experts at @nutanix :
Layered data pipelines, Bronze, Silver, Gold—
help teams manage raw, cleaned, and business-ready data efficiently.
Key for reliable analytics and scalable lakehouse operations.
#DataEngineering#ApacheIceberg
If Apache Iceberg adoption is fast, the ecosystem around it is growing even faster and that's creating one of the cheapest lakehouse infrastructure options available today.
I'm seeing some pretty exciting stuff happening in the cost-efficiency space.
If you're watching @ApacheIceberg adoption (and you should be), you've probably noticed the ecosystem around it is exploding even faster than the format itself. What's really caught my attention is how this is creating some of the cheapest lakehouse infrastructure options we've seen.
Here's the stack I'm seeing smart teams build:
🔹 @_olake for real-time replication
🔹 Apache Iceberg as the open table format
🔹 @ClickHouseDB for analytics (with some game-changing recent updates)
Let me break down why this combo is working so well...
OLake is doing the heavy lifting where it matters most - streaming data from your operational databases (@PostgreSQL , @MySQL, @MongoDB @) straight into Iceberg.
We're talking 46K+ records/second throughput here. And here's the kicker: it's open-source and doesn't need the usual suspects like ApacheSpark , Flink, or Debezium. Just clean, direct replication to all major Iceberg catalogs.
Iceberg brings the foundation - that open table format that eliminates vendor lock-in while giving you ACID transactions, schema evolution, and time travel. You know, all the stuff that used to be expensive and proprietary.
But here's where it gets interesting - ClickHouse just dropped some major updates in v25.8 that are making this stack even more compelling:
✅ Native Write Support - Full CRUD operations, not just reads anymore
✅ Production-Ready Catalogs - REST, Glue, Unity all promoted from experimental
✅ Schema Evolution - Add/drop/modify columns without breaking a sweat
✅ Better Deletes - Position deletes merged efficiently
✅ Near Real-time Streaming - Perfect match for ingestion platforms like OLake
Why this matters for your infrastructure costs:
Traditional warehouses are still charging premium prices for what this open stack delivers at a fraction of the cost.
ClickHouse alone is showing 5-15x cost advantages over traditional warehouses, and when you combine it with free, open-source ingestion and storage layers, the economics become pretty compelling.
The pattern I'm seeing:
Real-time ingestion → Open storage → Fast analytics = Maximum performance at minimum cost
Anyone else experimenting with similar stacks? Would love to hear what combinations are working (or not working) for you.
#ApacheIceberg #OpenTableFormats #ClickHouse #DataLakehouse #DataEngineering
@GoogleCloudTech Hey Google! We at OLake have been building a blazing fast ingestion tool to ease the adoption of Apache Iceberg. We support ingesting directly to the GCS with any REST Catalog.
Checkout our docs at: https://t.co/BOpAkJuMRp
@parmardarshil07 So we built out a blazing fast open-source ingestion tool to ease the adoption.
We recently released a helm chart for the same. Do check it out!!
Link: https://t.co/c79GWBpPbb
@parmardarshil07 Great explanation Darshil 👌
Iceberg is truly a great technology. It is the future!!
However the adoption for now is slow and not very approachable for smaller companies. We at OLake wanted to solve this very issue of adoption.
@theCUBE@DellTech@HPE@NFL Meanwhile we at OLake have been cooking too.
We recently released our OLake UI for Kubernetes to ease the adoption of Apa he Iceberg.
Blog link: https://t.co/c79GWBpPbb