TY - RPRT TI - Pre-trained Models Perform the Best When Token Distributions Follow Zipf's Law AU - Yanjin He AU - Qingkai Zeng AU - Meng Jiang PY - 2025 UR - https://arxiv.org/abs/2507.22543 ID - 2507.22543 ER -