Alex Merced's Data, Dev and AI Blog

Topic

Iceberg v4

7 posts tagged “Iceberg v4”.

  1. 32 min read

    What Iceberg v4's Proposed FILE Type Means for Multimodal Tables

    A product catalog table has two million rows. Each row has a SKU, a price, a category, and three product photos. The photos live in an object storage…

  2. 31 min read

    Parquet-Only Manifests in Iceberg v4: Why the Metadata Layer Is Going Columnar

    Picture a table with 40 million data files. Every one of those files has an entry in a manifest, and every entry carries per-column statistics for…

  3. 31 min read

    Iceberg v4's Adaptive Metadata Tree, Explained From First Principles

    The best way to understand the centerpiece of the Apache Iceberg v4 design effort is not to read the proposal first. It is to earn the proposal:…

  4. 30 min read

    Why Iceberg v4 Is Really About Making the Cost of Change Proportional to the Change

    Read enough of the Apache Iceberg v4 proposals, the design documents, the dev-list threads, the community sync notes, and a pattern emerges that no…

  5. 21 min read

    Reading the Apache Iceberg V4 Proposals Before They Land

    A Flink job commits every five seconds. Each commit writes one small Parquet file. It also writes a manifest, rewrites a manifest list, and writes a…

  6. 30 min read

    The State of Apache Iceberg v4 in July 2026: What the Dev List Tells Us About the Format's Next Chapter

    If you want to know where Apache Iceberg is headed, do not read the press releases. Read the dev mailing list. I say that as someone who reads it…

  7. 13 min read

    Iceberg v4 Performance: Root Manifests and Calls

    Apache Iceberg v4 discussion should focus on planning cost, metadata layout, and object storage round trips, not vague claims about faster tables.…

Browse all posts

Newsletter

Get new posts in your inbox

Deep dives on Apache Iceberg, lakehouse architecture and applied AI. No spam, unsubscribe anytime.

Subscribe

Menu

Search

Type at least two characters.