Abstract
Apache Iceberg V3 adoption is growing and the community is already shaping what comes next. This talk walks through the emerging scope and priorities for the V4 table spec and progress of the proposals currently underway.
We'll organize the roadmap into three themes. On metadata and scale, we'll cover the Adaptive Metadata Tree (AMT) and single-file commits, column statistics improvements, relative path support, and efficient column-level updates that reduce write amplification on wide tables. On table features, we'll look at default value expressions, generated columns, check constraints, collation support for string types, and virtual columns. And on new data types, we'll explore first-class file and vector types that open Iceberg to new workloads like AI and search.
Attendees will leave with a clear picture of where V4 is headed, why, which features are mature versus newly proposed, and how to get involved in shaping the spec before it lands.
$2,955, Conference (3 days). Current pricing ends October 13th. All pass options.
QCon San Francisco 2026 is a three day conference for senior software engineers, architects and team leads. An international program committee of working engineers selects every session. Patterns and practices, not products and pitches.
From the same track
Tuesday 17 November
10:35 Seacliff ABC Session Breaking Down the Walls to Python & Big Data + AI Performance with Transpilation and LLMs Holden Karau Data Engineer @Snowflake & Co-Founder of Fight Health Insurance, Previously @Netfilx, @Apple, and @Google Many of us have a love hate relationship with Apache Spark, especially when it comes to PySpark. Especially in Python, the UDF performance can be a dumpster fire. 11:45 Seacliff ABC Session Data Infrastructure Building a Data Platform with AI Mouli Mukherjee Engineering Manager @OpenAI Data platform sits at the intersection of infrastructure, governance, operations, and user experience. 13:35 Seacliff ABC Session One Does Not Simply Add Synchronous Writes - Bridging Batch and Real Time with Apache Iceberg Paige Elinson Staff Software Engineer @Datadog What happens when a batch ingestion platform needs to behave like an operational database? Datadog’s Reference Tables product was designed for asynchronous uploads of large datasets, but customers increasingly needed synchronous row-level edits, without sacrificing batch scale. 14:45 Seacliff ABC Session Building a StreamHouse: Turning Kafka Streams into a Millisecond Query Layer AI agents have compressed business decisions from minutes to seconds. Data systems need to keep up. 15:55 Seacliff D Unconference Unconference: Data Platforms Reimagined 17:05 Seacliff ABC Session Iceberg V4: The Next Chapter in Table Formats Apache Iceberg V3 adoption is growing and the community is already shaping what comes next. This talk walks through the emerging scope and priorities for the V4 table spec and progress of the proposals currently underway.