What Actually Happens When Two Engines Write the Same Iceberg Table at Once?
Somewhere in your platform, right now, a Spark job and a streaming writer are heading toward the same Apache Iceberg table, and they will…
The archive
593 posts on Apache Iceberg, lakehouse architecture, data engineering and applied AI.
Somewhere in your platform, right now, a Spark job and a streaming writer are heading toward the same Apache Iceberg table, and they will…
The Variant type in Apache Iceberg v3 gets described in one sentence so often that the sentence has started doing damage: "store JSON…
Two architects are arguing in a design review about whether to use "managed" Iceberg tables, and the argument is unresolvable, because they…
Here is a bill that surprises teams every quarter. A Flink pipeline streams change data capture events into an Apache Iceberg table, a few…
Every few years a data team discovers, mid-contract-renewal, exactly how much of their platform they do not control. The data sits in the…
Every architecture era gets its reference diagram. The warehouse era had its star schemas and its staging-to-mart flow. The big data era…
Every data platform has a quality system, and most of them are the same system: a few hundred rules, written after incidents, checking the…
The AI budget conversation changed its tone this year, and the numbers explain why. Industry surveys of enterprise AI spending in 2026 keep…
For twenty years, the cost of an ambiguous metric was a meeting. Two dashboards disagreed, two teams defended their numbers, someone…
Every article in my agentic lakehouse series has quietly assumed a plane that sees everything: the catalog's audit stream, the semantic…
Newsletter
Deep dives on Apache Iceberg, lakehouse architecture and applied AI. No spam, unsubscribe anytime.