TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Aug 3 – Aug 9, 2026
本周最热252

Recursive Synthesis for Long-Horizon Terminal Tasks

Zhongzhi Li, Yucheng Shi, Zongxia Li +8 authors

Recursive verified synthesis generates scalable long-horizon terminal-agent training data, substantially improving model performance on terminal benchmarks through supervised fine-tuning and reinforcement learning.

recursive verified synthesisterminal-agent taskssupervised fine-tuningagentic PPOHF ↗arXiv ↗

50 篇论文 · 按点赞排序

04

DAPD: Dual-Anchored Policy Distillation

Jianyu Wu, Yizhou Wang, Encheng Su +2 authors

Dual-Anchored Policy Distillation resolves privilege illusion in on-policy self-distillation by aligning matched-information paths and reducing reliance on privileged guidance, improving performance across model scales.

153on-policy self-distillationprivilege illusionHF ↗arXiv ↗
07

Mental World Modeling

Hao Fei, Yiran Zhao

Mental World Modeling integrates hidden mental states into predictive world models to improve action forecasting by jointly simulating physical scenes and agent beliefs.

110Mental World Modelingcoupled physical-mental world stateHF ↗arXiv ↗
12

WorldClaw: Agentic 3D Open-World Generation at Scale

Chunchao Guo, Jinpeng Li, Yang Li +1 authors

WorldClaw is an agentic coarse-to-fine framework that generates large-scale editable 3D worlds from text by combining planning agents, semantic layouts, reusable assets, and render-based refinement.

84coarse-to-fineplanning agentsHF ↗arXiv ↗
24

UEmbed: Unified Sparse and Dense Multimodal Embeddings

Tingyu Song, Mingxin Li, Yanzhao Zhang +5 authors

UEmbed is a decoder-only multimodal model that jointly produces dense and sparse embeddings in a single forward pass, extending sparse retrieval to unified text and multimodal inputs.

52learned sparse retrievaldecoder-onlyHF ↗arXiv ↗
27

GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?

Qifeng Zhang, Kaixiang Huang, Heng Dong +6 authors

The study introduces a video benchmark requiring global spatial reasoning across long videos and reveals that vision-language models struggle to build consistent global scene representations despite strong local perception.

46global spatial intelligenceVQA benchmarkHF ↗arXiv ↗
29

CADENA: Stepwise CAD Reverse Engineering

Soslan Kabisov, Gennadiy Savrasov, Maksim Elistratov +9 authors

CADENA reconstructs 3D meshes into parametric CAD programs step-by-step with intermediate geometry checks and introduces a benchmark for mechanical part reverse engineering.

42parametric CAD programreverse-engineeringHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号