TY - RPRT TI - LIFBench: Evaluating the Instruction Following Performance and Stability of Large Language Models in Long-Context Scenarios AU - Xiaodong Wu AU - Minhao Wang AU - Yichen Liu AU - Xiaoming Shi AU - He Yan AU - Xiangju Lu AU - Junmin Zhu AU - Wei Zhang PY - 2025 UR - https://arxiv.org/abs/2411.07037 ID - 2411.07037 ER -