TY - RPRT TI - Benchmarking Large Language Models on Floating-Point Error Classification AU - Lisa Taldir AU - Muhammad Ahmad Saeed AU - David Defour AU - Pablo de Oliveira Castro AU - Eric Petit PY - 2026 UR - https://arxiv.org/abs/2606.31308 ID - 2606.31308 ER -