arXiv · 2502.18771
Exploring Graph Learning Tasks with Pure LLMs: A Comprehensive Benchmark and Investigation
Abstract
In recent years, large language models (LLMs) have emerged as promising candidates for graph tasks. Many studies leverage natural language to describe graphs and apply LLMs for reasoning, yet most focus narrowly on performance benchmarks without fully comparing LLMs to graph learning models or exploring their broader potential. In this work, we present a comprehensive study of LLMs on graph learning tasks, evaluating both off-the-shelf and instruction-tuned models across a variety of scenarios. Beyond accuracy, we discuss data leakage concerns and computational overhead, and assess their performance under few-shot/zero-shot settings, domain transfer, structural understanding, and robustness. Our findings show that LLMs, particularly those with instruction tuning, greatly outperform traditional graph learning models in few-shot settings, exhibit strong domain transferability, and demonstrate excellent generalization and robustness. Our study highlights the broader capabilities of LLMs in graph learning and provides a foundation for future research.
Explore related subjects
Keep this discovery
Yuxiang Wang, Xinnan Dai, Wenqi Fan, Yao Ma. 2025-02-26. Exploring Graph Learning Tasks with Pure LLMs: A Comprehensive Benchmark and Investigation. https://arxiv.org/abs/2502.18771
Cite the original work for its findings. Save a collection to share your selection of sources.