TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Jul 7 – Jul 13, 2025

50 篇论文 · 按点赞排序

33

Critiques of World Models

Eric Xing, Mingkai Deng, Jinyu Hou +1 authors

A new architecture for a general-purpose world model is proposed, based on hierarchical, multi-level, and mixed continuous/discrete representations, with a focus on simulating actionable possibilities for purposeful reasoning and acting.

27world modelhierarchicalHF ↗arXiv ↗
35

GTA1: GUI Test-time Scaling Agent

Yan Yang, Dongxu Li, Yutong Dai +11 authors

A GUI Test-time Scaling Agent addresses task planning ambiguity and visual grounding accuracy in GUI interactions using reinforcement learning and test-time scaling.

27GUIaction proposalsHF ↗arXiv ↗
36

A Systematic Analysis of Hybrid Linear Attention

Dustin Wang, Rui-Jie Zhu, Steven Abreu +8 authors

Research evaluates various linear attention models in hybrid architectures, finding that selective gating, hierarchical recurrence, and controlled forgetting are crucial for effective recall in Transformers.

26transformersquadratic complexityHF ↗arXiv ↗
37

RoboBrain 2.0 Technical Report

BAAI RoboBrain Team, Mingyu Cao, Huajie Tan +43 authors

RoboBrain 2.0, a heterogeneous vision-language model, excels in embodied reasoning tasks with strong performance on spatial and temporal benchmarks, supporting capabilities like spatial understanding and temporal decision-making.

26embodied vision-language foundation modelsheterogeneous architectureHF ↗arXiv ↗
40

First Return, Entropy-Eliciting Explore

Tianyu Zheng, Tianshun Xing, Qingshui Gu +10 authors

FR3E, a structured exploration framework, enhances LLM reasoning by providing targeted guidance at high-uncertainty decision points, leading to more stable training and accurate responses.

24Reinforcement Learning from Verifiable RewardsLarge Language ModelsHF ↗arXiv ↗
45

Is Diversity All You Need for Scalable Robotic Manipulation?

Modi Shi, Li Chen, Jin Chen +7 authors

Investigation into data diversity in robotic manipulation reveals that task diversity is more critical than demonstration quantity, single-embodiment data can suffice for cross-embodiment transfer, and expert diversity can be confounding, leading to a distribution debiasing method that improves performance.

20data diversitytask diversityHF ↗arXiv ↗
46

Token Bottleneck: One Token to Remember Dynamics

Taekyung Kim, Dongyoon Han, Byeongho Heo +2 authors

ToBo, a self-supervised learning pipeline, generates compact and temporally aware visual representations by encoding scenes into a bottleneck token and predicting subsequent scenes with minimal hints, demonstrating superior performance in sequential tasks.

19Token Bottleneckself-supervised learningHF ↗arXiv ↗
47

Differential Mamba

Nadav Schneider, Itamar Zimerman, Eliya Nachmani

A novel differential mechanism for Mamba, a selective state-space layer architecture, improves language modeling performance by addressing overallocation of attention to irrelevant context.

19TransformersRNNsHF ↗arXiv ↗
2 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号