arXiv · 2511.13726
Refine Thought: A Test-Time Inference Method for Embedding Model Reasoning
Abstract
We propose RT (Refine Thought), a method that can enhance the semantic reasoning ability of text embedding models. The method obtains the final semantic representation by running multiple forward passes of the text embedding model. Experiments show that RT achieves significant improvements on semantic reasoning tasks in BRIGHT and the person-job matching benchmark PJBenchmark, while maintaining consistent performance on general-purpose semantic understanding tasks such as C-MTEB. Our results indicate that RT is effective because it further activates the semantic reasoning ability learned during pretraining by decoder-only text embedding models (e.g., Qwen3-Embedding-8B). RT can be seen as a test-time inference method.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Guangzhi Wang, Kai Li, Yinghao Jiao, Zhi Liu. 2025-10-14. Refine Thought: A Test-Time Inference Method for Embedding Model Reasoning. https://arxiv.org/abs/2511.13726
Cite the original work for its findings. Save a collection to share your selection of sources.