ReportGem ReportGem

Academic paper

AI Query Compilation for Unified and Optimized Execution

Authors: Yeounoh Chung, Helena Caminal, Fatma OzcanPublished: 2026-08-10Paper ID: 2608.10139Category: cs.DBLicense: CC BY 4.0

Abstract

In this vision paper, we propose a novel architectural paradigm for accelerated AI query execution via a unified compiled execution strategy. By compiling the hybrid AI Query as a whole -- integrating both standard SQL relational constructs and LLM inference layers into a single, unified tensor compute graph -- we completely alleviate PCIe data movement bottlenecks across execution boundaries and enable global compiler optimizations and efficient automatic sharding. We demonstrate the viability of this unified execution paradigm on select and extended AI queries on SemBench Reviews and Movies datasets, achieving up to 5.3x latency speedup and 9.8x throughput speedup on TPUs, and outline a research roadmap of open technical challenges to realize this vision.

This public page contains bibliographic metadata and the author abstract. Use the reader for licensed document access.

Open licensed paper reader