SearcharxivSearch

arXiv subjects

Arush Garg

Publications and source records attributed to Arush Garg.

2 recordsLinked to original sources

Know Your Limits : On the Faithfulness of LLMs as Solvers and Autoformalizers in Legal Reasoning

Large Language Models (LLMs) achieve strong performance on reasoning tasks, but whether this reflects faithful logical inference or heuristic approximation remains unclear. We study this question in legal entailment by comparing three paradigms, including pure LLM classification, LLM-based Formal Reasoning, and solver-based Formal Reasoning using the Z3 SMT solver, on a re-annotated subset of ContractNLI across five LLMs. Our re-annotation reveals a systematic and measurable gap between pragmatic legal interpretation and strict formal entailment, where a substantial proportion of legally sound inferences are not formally grounded without additional unstated assumptions. While introducing formal structure improves accuracy, with LLM-based Formal Reasoning achieving the highest benchmark performance, we show that this gain does not imply faithful reasoning. We identify three recurring failure modes: scope laundering, where LLMs report solver-inconsistent classifications without executing the underlying formal reasoning, producing conclusions that appear logically grounded but are not; implicit constraint blindness, where LLMs overlook logical constraints present in formal representations; and program synthesis failures, where LLMs generate incorrect Z3 code despite structured prompting. Critically, scope laundering persists across all models, raising serious concerns about the faithfulness of LLM-based formal reasoning as a proxy for symbolic execution. These results reveal a fundamental gap between benchmark accuracy and logical faithfulness.

cs.AI

Assessing the dynamical assumptions in Tsirelson inequality tests of non-classicality in harmonic oscillators

"Macrorealism" posits that a system possesses definite properties at all times and that we can discover these properties, in principle, without disturbing the system's subsequent behaviour. The Leggett-Garg inequalities are derived under these assumptions and are readily violated by standard quantum mechanics, thereby providing a scheme to test whether demonstrably macroscopic systems can exhibit quantum coherence. Unfortunately, Leggett-Garg tests suffer from the difficult to avoid clumsiness loophole - the difficulty of proving that sequential measurements have not inadvertently disturbed the system. The recently uncovered Tsirelson inequality is derived from the simple dynamical assumption of uniform precession, obeyed by many classical systems, and requires only single-time measurements. However, Tsirelson inequality violations could be explained by a macrorealistic system that merely breaks the dynamical assumption, rather than genuine quantum behaviour. By carrying out a quantum-mechanical analysis of the Tsirelson inequality in the harmonic oscillator, we develop a protocol to rule out this possibility by assessing generalised conditions of uniform precession. We show that various measures of uniform precession, some of which are related to Leggett-Garg quantities, are satisfied well enough that the presence of quantum-mechanical interference terms must be implied. We derive several incidental mathematical results relating to violating states of Tsirelson's inequality, concerning dwell time, crossing number and probability currents, and also consider a group theoretic analysis of the Tsirelson operator.

quant-ph