TensorX
返回文献探索

Paper · arXiv 2606.27378

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

Fahd Seddik, Fatemeh Fard

61 upvotesMay 7, 2026arXiv 预印本
AI 摘要

An axiomatic evaluation framework reveals systematic failures in latent thought representations of LLMs across multiple reasoning tasks, demonstrating that current representations fail to satisfy fundamental functional axioms consistently across different model architectures.

axiomatic evaluation frameworklatent thought representationsLLMsfunctional axiomscausalityminimalityseparabilitystabilitydownstream benchmark scoresrepresentation qualitymodel capacityreasoning tasksspatial reasoningfactual QAopen-weight LLMsrepresentation failurestructural gap

Abstract

We introduce an axiomatic evaluation framework for latent thought representations in LLMs, comprising metrics that are independent of downstream benchmark scores and reveal representational failures that benchmark accuracy masks. Existing evaluations conflate representation quality with model capacity. Therefore, failures cannot be attributed to the representation rather than to the model that processes it. We formalize four functional axioms (Causality, Minimality, Separability, and Stability) and define a quantitative measure for each, computed directly on the representation independently of downstream accuracy. We audit open-weight LLMs across 23 reasoning tasks (e.g., Spatial Reasoning, Factual QA). We find that no candidate satisfies all four axioms simultaneously, that the representations distinguish task type reliably but cannot distinguish between two questions within the same task, and that the representations encode little information beyond what is already present in the input embedding. The failure is consistent across dense, reasoning-distilled, and RL-trained model families, indicating that the gap is structural rather than a property of model size or training procedure.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs | TensorX