TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Oct 6 – Oct 12, 2025
本周最热521

Less is More: Recursive Reasoning with Tiny Networks

Alexia Jolicoeur-Martineau

Tiny Recursive Model (TRM) achieves high generalization on complex puzzle tasks using a small, two-layer network with minimal parameters, outperforming larger language models.

Hierarchical Reasoning ModelHRMTiny Recursive ModelTRMHF ↗arXiv ↗

50 篇论文 · 按点赞排序

02

Agent Learning via Early Experience

Kai Zhang, Xiangchao Chen, Bo Liu +27 authors

Early experience, using agent-generated interaction data without reward signals, improves policy effectiveness and generalization, serving as a bridge between imitation learning and reinforcement learning.

276reinforcement learningearly experienceHF ↗arXiv ↗
04

Apriel-1.5-15b-Thinker

Shruthan Radhakrishna, Aman Tiwari, Aanjaneya Shukla +21 authors

A 15-billion parameter multimodal reasoning model achieves competitive performance through a progressive training methodology without reinforcement learning, demonstrating efficient use of computational resources.

125multimodal reasoning modeldepth upscalingHF ↗arXiv ↗
05

Paper2Video: Automatic Video Generation from Scientific Papers

Zeyu Zhu, Kevin Qinghong Lin, Mike Zheng Shou

PaperTalker is a multi-agent framework that automates academic presentation video generation by integrating slide generation, layout refinement, subtitling, speech synthesis, and talking-head rendering, outperforming existing methods.

120multi-agent frameworkslide generationHF ↗arXiv ↗
07

MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Policy Optimization

Xiangyu Zhao, Junming Lin, Tianhao Liang +11 authors

Existing Multimodal Large Language Models show performance deficits in long-chain reflective reasoning, which is addressed by developing MM-HELIX-100K and Adaptive Hybrid Policy Optimization, leading to improved accuracy and generalization.

110Multimodal Large Language Modelslong-chain reflective reasoningHF ↗arXiv ↗
09

UniVideo: Unified Understanding, Generation, and Editing for Videos

Cong Wei, Quande Liu, Zixuan Ye +5 authors

UniVideo, a dual-stream framework combining a Multimodal Large Language Model and a Multimodal DiT, extends unified modeling to video generation and editing, achieving state-of-the-art performance and supporting task composition and generalization.

81Multimodal Large Language ModelMultimodal DiTHF ↗arXiv ↗
13

DreamOmni2: Multimodal Instruction-based Editing and Generation

Bin Xia, Bohao Peng, Yuechen Zhang +10 authors

DreamOmni2 addresses limitations in instruction-based image editing and subject-driven generation by introducing multimodal instruction-based editing and generation tasks, utilizing feature mixing, index encoding, and joint training with a VLM.

74instruction-based editingsubject-driven generationHF ↗arXiv ↗
14

MemMamba: Rethinking Memory Patterns in State Space Model

Youjin Wang, Yangjingyi Chen, Jiahao Yan +2 authors

MemMamba, a novel architecture integrating state summarization and cross-attention, improves long-range memory and efficiency in sequence modeling compared to Mamba and Transformers.

74recurrent neural networksgradient vanishingHF ↗arXiv ↗
18

Fast-dLLM v2: Efficient Block-Diffusion LLM

Chengyue Wu, Hao Zhang, Shuchen Xue +7 authors

Fast-dLLM v2, a block diffusion language model, efficiently converts pretrained autoregressive models for parallel text generation, achieving significant speedup without compromising accuracy.

60autoregressive modelslarge language modelsHF ↗arXiv ↗
24

Training-Free Group Relative Policy Optimization

Yuzheng Cai, Siqi Cai, Yuchen Shi +10 authors

Training-Free GRPO enhances LLM agent performance in specialized domains by learning experiential knowledge as a token prior without parameter updates, improving out-of-domain tasks with minimal data.

46Large Language Model (LLM)agentic reinforcement learningHF ↗arXiv ↗
26

CoDA: Coding LM via Diffusion Adaptation

Haolin Chen, Shiyu Wang, Can Qin +12 authors

CoDA, a 1.7B-parameter diffusion coder, achieves competitive performance with smaller models through confidence-guided sampling and is released with open-source tools.

43diffusion language modelsbidirectional contextHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号