Papers
Topics
Authors
Recent
Search
2000 character limit reached

Beyond Confidence: The Rhythms of Reasoning in Generative Models

Published 11 Feb 2026 in cs.CL and cs.AI | (2602.10816v1)

Abstract: LLMs exhibit impressive capabilities yet suffer from sensitivity to slight input context variations, hampering reliability. Conventional metrics like accuracy and perplexity fail to assess local prediction robustness, as normalized output probabilities can obscure the underlying resilience of an LLM's internal state to perturbations. We introduce the Token Constraint Bound (δ<em>TCBδ<em>{\mathrm{TCB}}), a novel metric that quantifies the maximum internal state perturbation an LLM can withstand before its dominant next-token prediction significantly changes. Intrinsically linked to output embedding space geometry, δ</em>TCBδ</em>{\mathrm{TCB}} provides insights into the stability of the model's internal predictive commitment. Our experiments show δ<em>TCBδ<em>{\mathrm{TCB}} correlates with effective prompt engineering and uncovers critical prediction instabilities missed by perplexity during in-context learning and text generation. δ</em>TCBδ</em>{\mathrm{TCB}} offers a principled, complementary approach to analyze and potentially improve the contextual stability of LLM predictions.

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.