TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Jun 29 – Jul 5, 2026
本周最热508

Orca: The World is in Your Mind

Yihao Wang, Yuheng Ji, Mingyu Cao +54 authors

Orca establishes a unified world latent space through next-state-prediction modeling using multimodal data and demonstrates superior performance in downstream tasks compared to specialized baselines.

world foundation modelworld latent spacemultimodal readout interfacesnext-state-prediction modelingHF ↗arXiv ↗

50 篇论文 · 按点赞排序

02

Program-as-Weights: A Programming Paradigm for Fuzzy Functions

Wentao Zhang, Liliana Hotsko, Woojeong Kim +3 authors

Fuzzy-function programming compiles natural-language specifications into compact neural artifacts using a 4B compiler and 0.6B interpreter, achieving efficient, local execution with reduced memory usage and faster inference.

310Program-as-WeightsFuzzyBenchHF ↗arXiv ↗
05

DOPD: Dual On-policy Distillation

Xinlei Yu, Gen Li, Qingyi Si +13 authors

DOPD addresses privilege illusion in on-policy distillation by dynamically routing token-level supervision between teacher and student policies based on advantage gaps and probabilities, improving capability transfer in large and vision-language models.

116on-policy distillationtoken-level signalsHF ↗arXiv ↗
06

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

Lei Bai, Zongsheng Cao, Yang Chen +47 authors

Agents-A1, a 35B Mixture-of-Experts Agentic Model, achieves trillion-parameter-level performance through long-horizon trajectory scaling and heterogeneous agent ability scaling via a three-stage training approach involving supervised fine-tuning, domain-level teacher models, and multi-teacher distillation.

106Mixture-of-Expertsagentic modelHF ↗arXiv ↗
10

AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents

Xiangchen Cheng, Yunwei Jiang, Jianwen Sun +7 authors

A bounded contract approach for long-horizon LLM agents uses typed retrieval to assemble fresh prompts, enabling isolated analysis of memory components and demonstrating improved performance in complex decision-making tasks.

71long-horizon LLM agentbounded contractHF ↗arXiv ↗
11

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

Fahd Seddik, Fatemeh Fard

An axiomatic evaluation framework reveals systematic failures in latent thought representations of LLMs across multiple reasoning tasks, demonstrating that current representations fail to satisfy fundamental functional axioms consistently across different model architectures.

61axiomatic evaluation frameworklatent thought representationsHF ↗arXiv ↗
12

Qwen-Image-2.0-RL Technical Report

Yixian Xu, Kaiyuan Gao, Yuxiang Chen +25 authors

A reinforcement learning and on-policy distillation approach enhances the visual quality and instruction-following capabilities of a diffusion model for image generation and editing tasks.

55reinforcement learning from human feedbackon-policy distillationHF ↗arXiv ↗
15

Morphing into Hybrid Attention Models

Disen Lan, Jianbin Zheng, Yuxi Ren +5 authors

FlashMorph is an efficient layer selection method that formulates hybrid layer selection as a budget-constrained optimization problem, using morphable models and linearization regularization to improve long-context efficiency in Transformers.

53hybrid attention modelsfull-attention layersHF ↗arXiv ↗
17

TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents

Shoufa Chen, Luyuan Wang, Xuan Yang +7 authors

TUA-Bench presents a comprehensive benchmark for evaluating general-purpose terminal-use agents across diverse digital activities and specialized workflows, revealing significant performance gaps among current frontier agents.

49terminal-use agentsgeneral-purpose agentsHF ↗arXiv ↗
18

ReFreeKV: Towards Threshold-Free KV Cache Compression

Xuanfan Ni, Liyan Xu, Chenyang Lyu +6 authors

ReFreeKV addresses the limitations of threshold-dependent KV cache pruning by introducing a threshold-free approach that adaptively allocates compression budgets while maintaining full-cache performance across diverse datasets and model sizes.

48KV cache pruningthreshold-free methodsHF ↗arXiv ↗
20

Beyond IID: How General Are Tabular Foundation Models, Really?

Lennart Purucker, Andrej Tschalzev, Nick Erickson +7 authors

Tabular foundation models show varying performance across different data conditions, with traditional methods still outperforming newer approaches on complex, large-scale datasets.

44tabular foundation modelspredictive machine learningHF ↗arXiv ↗
22

Trimming the Long-Tail of Visual World Modeling Evaluation

Bingxuan Li, Yining Hong, Cheng Qian +6 authors

Current visual world models demonstrate limited generalization beyond common physical interactions, struggling with rare and irregular scenarios despite achieving realism on standard benchmarks.

44visual world modelsphysical interactionsHF ↗arXiv ↗
23

Multi-Block Diffusion Language Models

Yijie Jin, Jiajun Xu, Yuxuan Liu +8 authors

Multi-Block Diffusion Language Models extend single-block diffusion to concurrent block decoding with improved training strategies and optimized decoding algorithms.

43Block Diffusion Language ModelsMulti-Block DiffusionHF ↗arXiv ↗
26

Translation as a Bridging Action: Transferring Manipulation Skills from Humans to Robots

Sijin Chen, Kaixuan Jiang, Haixin Shi +6 authors

Human manipulation skills are transferred to robots more effectively by using a bridging action representation based on relative wrist translation in the initial head-camera frame, combined with a vision-language-action model that handles embodiment differences through interleaved action tokens and attention masking.

41bi-manual manipulationparallel grippersHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号