TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Sep 15 – Sep 21, 2025
本周最热119

FlowRL: Matching Reward Distributions for LLM Reasoning

Xuekai Zhu, Daixuan Cheng, Dinghuai Zhang +20 authors

FlowRL enhances LLM reinforcement learning by matching the full reward distribution through flow balancing, improving diversity and performance over reward-maximizing methods.

FlowRLreward distributionflow balancingreinforcement learningHF ↗arXiv ↗

50 篇论文 · 按点赞排序

02

Scaling Agents via Continual Pre-training

Liangcai Su, Zhen Zhang, Guangyu Li +19 authors

AgentFounder, a deep research agent model incorporating Agentic Continual Pre-training, achieves state-of-the-art performance in agentic tasks while maintaining strong tool-use ability.

118Large language modelsagentic systemsHF ↗arXiv ↗
13

SAIL-VL2 Technical Report

Weijie Yin, Yongjie Ye, Fangxun Shu +11 authors

SAIL-VL2, a vision-language foundation model, achieves state-of-the-art performance across diverse benchmarks through data curation, progressive training, and sparse MoE architecture.

46vision-language foundation modelSAIL-VL2HF ↗arXiv ↗
15

AToken: A Unified Tokenizer for Vision

Jiasen Lu, Liangchen Song, Mingze Xu +5 authors

AToken, a unified visual tokenizer, achieves high-fidelity reconstruction and semantic understanding across images, videos, and 3D assets using a 4D transformer architecture with adversarial-free training.

37unified visual tokenizer4D latent spaceHF ↗arXiv ↗
17

Single-stream Policy Optimization

Zhongwen Xu, Zihan Ding

Single-stream Policy Optimization (SPO) improves policy-gradient training for Large Language Models by eliminating group-based issues and providing a stable, low-variance learning signal, leading to better performance and efficiency.

36policy-gradient optimizationLarge Language Models (LLMs)HF ↗arXiv ↗
25

Lost in Embeddings: Information Loss in Vision-Language Models

Wenyan Li, Raphael Tang, Chengzu Li +3 authors

Two approaches are introduced to analyze and quantify information loss in vision-language models during the projection of visual inputs into the language model's embedding space, revealing significant distortions and their impact on model performance.

29vision--language modelspretrained vision encoderHF ↗arXiv ↗
26

PANORAMA: The Rise of Omnidirectional Vision in the Embodied AI Era

Xu Zheng, Chenfei Liao, Ziqiao Weng +12 authors

Recent advancements in omnidirectional vision, driven by industrial and academic interest, have led to breakthroughs in generation, perception, and understanding, with the proposal of a new system architecture called PANORAMA.

28omnidirectional vision360-degree visionHF ↗arXiv ↗
28

Virtual Agent Economies

Nenad Tomasev, Matija Franklin, Joel Z. Leibo +4 authors

The sandbox economy framework analyzes the emerging AI agent economy, focusing on its origins and permeability, and discusses design choices for safe and steerable AI markets.

27HF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号