SearcharxivSearch

arXiv subjects

Akira Okutomi

Publications and source records attributed to Akira Okutomi.

3 recordsLinked to original sources

Stable Miscalibration in Large Language Models: A Practical View of High-Confidence Errors

High-confidence errors in large language models are often treated as evidence of fragile internal inference. We study a different possibility: stable miscalibration, where a confident wrong answer remains locally stable under small perturbations. We combine two diagnostics: a label-aware output-level audit score that ranks domains by confidence variation and overconfident mistakes under a forced-answer baseline, and an internal sensitivity probe that measures hidden-state movement. On a multi-domain binary factual audit set, this audit score tracks where abstention-aware self-critique reduces decision loss, although direct labeled baselines rank the same gain more strongly. Internally, self-critical prompting consistently reduces hidden-state sensitivity across layers in three open-weight models. This supports prompt-induced local stabilization rather than a purely output-level abstention pattern, but it does not imply calibration: audit-defined overconfident errors are not clearly more locally sensitive than confidently correct answers, so some high-confidence errors may be stable and miscalibrated rather than simply fragile.

cs.AI

False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs

High-confidence errors in large language models are often treated as fragile failures. We study an alternative: some errors may be false fixed points, locally stable, internally coherent, and confidently wrong. This separates robustness from truth-tracking. We develop the separation through a Kantian commitment-gate framing and a minimal linear feedback model in which stability and correctness can diverge. Across three open-weight models, overconfident wrong items are not systematically more locally fragile than confidently correct items under our hidden-state sensitivity probes. Abstention-aware self-critique reduces overconfident wrong commitments by sacrificing coverage, and C3-R, a rule-based explicit feedback gate, sharpens that tradeoff rather than eliminating it. These results motivate, but do not establish, high signal-to-noise (high-SNR) inertia and representational compression as possible mechanisms for stable miscalibration.

cs.AI

Conduction Effect of Thermal Radiation in a Metal Shield Pipe in a Cryostat for a Cryogenic Interferometric Gravitational Wave Detector

A large heat load caused by thermal radiation through a metal shield pipe was observed in a cooling test of a cryostat for a prototype of a cryogenic interferometric gravitational wave detector. The heat load was approximately 1000 times larger than the value calculated by the Stefan-Boltzmann law. We studied this phenomenon by simulation and experiment and found that it was caused by the conduction of thermal radiation in a metal shield pipe.

physics.ins-det