arXiv · 2306.00976
TopEx: Topic-based Explanations for Model Comparison
Abstract
Meaningfully comparing language models is challenging with current explanation methods. Current explanations are overwhelming for humans due to large vocabularies or incomparable across models. We present TopEx, an explanation method that enables a level playing field for comparing language models via model-agnostic topics. We demonstrate how TopEx can identify similarities and differences between DistilRoBERTa and GPT-2 on a variety of NLP tasks.
Explore related subjects
Keep this discovery
Shreya Havaldar, Adam Stein, Eric Wong, Lyle Ungar. 2023-06-01. TopEx: Topic-based Explanations for Model Comparison. https://arxiv.org/abs/2306.00976
Cite the original work for its findings. Save a collection to share your selection of sources.