← Tracks

Data + AI

Track Chairs
Juan Pan Sheng Wu Jeff Feng
Sessions
12
Room
JingMing Hall

In the era of Generative AI and Autonomous Agents, data is no longer just a static asset—it is the dynamic fuel for intelligence. The Data + AI track explores the profound convergence where Apache’s battle-tested data infrastructure meets cutting-edge AI capabilities.

We focus on two critical evolutionary paths:

  • Infrastructure for AI: How Apache projects are evolving to support the massive scale required by LLMs—covering vector retrieval, real-time RAG pipelines, and unstructured data governance.

  • AI-Evolved Data Systems: How AI agents and models are redefining data engineering itself—from autonomous query optimization and self-healing pipelines to AI-native features within established projects.

This track is designed for engineers and architects who move beyond the hype. Join us to dissect concrete architectures, production-grade agentic workflows, and the future of open standards in building reliable, governed, and scalable intelligent systems.

Agenda

2026-08-07 Full schedule →
  1. 14:00GMT+8
    From Lakehouse to Multimodal Data Lake: Rethinking Data Infrastructure for AI
    Zheng Yubin, Lili Ma
    JingMing Hall Chinese Session
  2. 14:30GMT+8
    Ant Real-Time Computing Team's Practice on Flink Expert Agent
    Chaoming Zhang
    JingMing Hall Chinese Session
  3. 15:00GMT+8
  4. 15:45GMT+8
  5. 16:15GMT+8
    Building a Palantir-like Data & AI Platform with the Apache Stack
    Zhangjian He
    JingMing Hall Chinese Session
  6. 16:45GMT+8