The paper “How Language Models Fail” studies reasoning failures in language models through token-level uncertainty signals. It distinguishes committed failures, where a model locks onto a wrong reasoning path early, from persistent uncertainty, where uncertainty accumulates throughout the trace and the full reasoning sequence is needed for diagnosis.