TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Aug 11 – Aug 17, 2025
本周最热215

GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models

GLM-4. 5 Team, Aohan Zeng, Xin Lv +168 authors

GLM-4.5, a Mixture-of-Experts large language model with 355B parameters, achieves strong performance across agentic, reasoning, and coding tasks using multi-stage training and reinforcement learning.

Mixture-of-Expertshybrid reasoning methodmulti-stage trainingexpert model iterationHF ↗arXiv ↗

50 篇论文 · 按点赞排序

06

WideSearch: Benchmarking Agentic Broad Info-Seeking

Ryan Wong, Jiawei Wang, Junjie Zhao +10 authors

WideSearch is a new benchmark evaluating the reliability of automated search agents in large-scale information collection tasks, revealing significant deficiencies in current systems.

113Large Language Modelsautomated search agentsHF ↗arXiv ↗
16

Part I: Tricks or Traps? A Deep Dive into RL for LLM Reasoning

Zihe Liu, Jiashun Liu, Yancheng He +12 authors

A systematic review of reinforcement learning techniques for large language model reasoning reveals clear guidelines and demonstrates that a minimalist combination of techniques can improve performance over existing strategies.

50reinforcement learningLLM reasoningHF ↗arXiv ↗
21

Klear-Reasoner: Advancing Reasoning Capability via Gradient-Preserving Clipping Policy Optimization

Zhenpeng Su, Leiyu Pan, Xue Bai +5 authors

Klear-Reasoner, a model with long reasoning capabilities, achieves high performance across benchmarks through detailed post-training workflows, including long Chain-of-Thought supervised fine-tuning and reinforcement learning with Gradient-Preserving clipping Policy Optimization.

43long reasoningChain-of-Thought supervised fine-tuningHF ↗arXiv ↗
23

Complex Logical Instruction Generation

Mian Zhang, Shujian Liu, Sixun Dong +9 authors

LogicIFGen and LogicIFEval assess the instruction-following capabilities of LLMs on complex, logic-rich instructions, revealing significant performance gaps.

40Large Language ModelsLLMsHF ↗arXiv ↗
28

Memp: Exploring Agent Procedural Memory

Runnan Fang, Yuan Liang, Xiaobin Wang +6 authors

Agents equipped with a learnable, updatable procedural memory system, Memp, achieve improved performance and efficiency across tasks by distilling past experiences into detailed instructions and higher-level abstractions.

36Large Language Modelsprocedural memoryHF ↗arXiv ↗
29

A Survey on Diffusion Language Models

Tianyi Li, Mingda Chen, Bowei Guo +1 authors

Diffusion Language Models offer parallel token generation, reducing inference latency and capturing bidirectional context, and are compared to autoregressive models in various NLP tasks.

35Diffusion Language ModelsautoregressiveHF ↗arXiv ↗
30

Puppeteer: Rig and Animate Your 3D Models

Chaoyue Song, Xiu Li, Fan Yang +6 authors

Puppeteer is a framework that automates rigging and animation of 3D models using an auto-regressive transformer, attention-based architecture, and differentiable optimization, outperforming existing methods in accuracy and efficiency.

33auto-regressive transformerjoint-based tokenizationHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号