arXiv · 2604.18964
DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning
Abstract
This paper introduces DW-Bench, a new benchmark that evaluates large language models (LLMs) on graph-topology reasoning over data warehouse schemas, explicitly integrating both foreign-key (FK) and data-lineage edges. The benchmark comprises 1,046 automatically generated, verifiably correct questions across five schemas. Experiments show that tool-augmented methods substantially outperform static approaches but plateau on hard compositional subtypes.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ahmed G. A. H Ahmed, C. Okan Sakar. 2026-04-21. DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning. https://arxiv.org/abs/2604.18964
Cite the original work for its findings. Save a collection to share your selection of sources.