Iceberg UDF Spec: Portable SQL Functions Across Engines
Huaxin Gao
English Session 2026-08-09 13:30 GMT+8 (ROOM : WanChun Hall) #datalakeIceberg’s UDF spec introduces a self-contained, versioned metadata format for SQL scalar and table functions that can move across catalogs and engines like Spark and Trino. This talk walks through the core model, definitions, parameter naming rules, return types, and dialect-specific representations, and explains how versioning, rollback, determinism, and null-handling hints make UDFs portable without sacrificing engine optimizations. We’ll highlight key design choices such as immutable signatures, overload compatibility with defaults, and secure functions, and close with practical guidance for implementing the spec in an engine or catalog.
Speakers:

Huaxin Gao: Software engineer at Snowflake
Huaxin Gao is a software engineer at Snowflake and an Apache Spark committer and PMC member. She is also a committer for Apache Iceberg and Apache DataFusion Comet, with contributions spanning query engines, table formats, and distributed data systems.