TensorX
返回文献探索

Paper · arXiv 2505.21189

Exploring the Latent Capacity of LLMs for One-Step Text Generation

Gleb Mezentsev, Ivan Oseledets

62 upvotesMay 27, 2025arXiv 预印本
AI 摘要

LLMs can generate long text segments in a single forward pass using learned embeddings, revealing a capability for multi-token generation without iterative decoding.

large language modelsautoregressive generationinput embeddingfrozen LLMsmulti-token generationiterative decodinglearned embeddingsembedding spacededicated encoder

Abstract

A recent study showed that large language models (LLMs) can reconstruct surprisingly long texts - up to thousands of tokens - via autoregressive generation from just one specially trained input embedding. In this work, we explore whether such reconstruction is possible without autoregression. We show that frozen LLMs can generate hundreds of accurate tokens in just one forward pass, when provided with only two learned embeddings. This reveals a surprising and underexplored capability of LLMs - multi-token generation without iterative decoding. We investigate the behaviour of these embeddings and provide insight into the type of information they encode. We also empirically show that although these representations are not unique for a given text, they form connected and local regions in embedding space - a property that suggests the potential of learning a dedicated encoder into that space.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Exploring the Latent Capacity of LLMs for One-Step Text Generation | TensorX