TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Jul 21 – Jul 27, 2025

50 篇论文 · 按点赞排序

31

CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models

Quang-Binh Nguyen, Minh Luu, Quang Nguyen +2 authors

CSD-VAR, a Visual Autoregressive Modeling approach, enhances content-style decomposition by introducing scale-aware optimization, SVD-based rectification, and augmented K-V memory, outperforming diffusion models in content preservation and stylization.

26content-style decompositionVisual Autoregressive ModelingHF ↗arXiv ↗
40

Inverse Scaling in Test-Time Compute

Aryo Pradipta Gema, Alexander Hägele, Runjin Chen +11 authors

Evaluation of Large Reasoning Models across different reasoning lengths reveals that increased test-time compute can lead to performance degradation and amplify problematic reasoning patterns.

21Large Reasoning Modelsreasoning lengthHF ↗arXiv ↗
42

Stabilizing Knowledge, Promoting Reasoning: Dual-Token Constraints for RLVR

Jiakang Wang, Runze Liu, Fuzheng Zhang +2 authors

Archer, an entropy-aware RLVR approach with dual-token constraints and synchronous updates, enhances LLM reasoning abilities by differentiating between knowledge and reasoning tokens, achieving state-of-the-art performance on mathematical reasoning and code generation benchmarks.

21Reinforcement Learning with Verifiable Rewards (RLVR)Large Language Models (LLMs)HF ↗arXiv ↗
48

Streaming 4D Visual Geometry Transformer

Dong Zhuo, Wenzhao Zheng, Jiahe Guo +3 authors

A streaming 4D visual geometry transformer uses causal attention and knowledge distillation to achieve real-time 4D reconstruction with high spatial consistency and competitive performance.

154D spatial-temporal geometrystreaming 4D visual geometry transformerHF ↗arXiv ↗
49

Hierarchical Budget Policy Optimization for Adaptive Reasoning

Shangke Lyu, Linjuan Wu, Yuchen Yan +7 authors

Hierarchical Budget Policy Optimization (HBPO) is a reinforcement learning framework that optimizes reasoning depth for large models, improving efficiency and accuracy by adapting to problem complexity.

14Hierarchical Budget Policy OptimizationHBPOHF ↗arXiv ↗
2 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号