TensorX

Trends · 研究趋势

数据来自 Hugging Face 论文的 AI 提取关键词,按月统计研究方向的增长与热度。

返回趋势

large reasoning models 相关论文

11 篇论文 · 按点赞排序

01

Metis: Memory Foundation Model

Zeyu Zhang, Ziliang Guo, Yihang Sun +14 authors

Metis introduces memory foundation models that embed persistent, dynamically evolving native memory states and autonomous storage procedures directly into foundation models via memory attention and gradient-free updates.

272memory foundation modelsnative memoryHF ↗arXiv ↗
03

Efficient Reasoning with Balanced Thinking

Yulin Li, Tengyao Tu, Li Ding +5 authors

ReBalance is a training-free framework that balances reasoning in large models by using confidence indicators to detect and correct overthinking and underthinking behaviors through dynamic steering vectors.

151large reasoning modelsoverthinkingHF ↗arXiv ↗
06

START: Self-taught Reasoner with Tools

Chengpeng Li, Mingfeng Xue, Zhenru Zhang +7 authors

START integrates external tools into large reasoning models to enhance capabilities, using techniques like Hint-infer and Hint Rejection Sampling Fine-Tuning, achieving high performance across various benchmarks.

113Large reasoning modelsChain-of-thoughtHF ↗arXiv ↗
07

Search-o1: Agentic Search-Enhanced Large Reasoning Models

Xiaoxi Li, Guanting Dong, Jiajie Jin +5 authors

Search-o1 enhances large reasoning models with an agentic retrieval-augmented generation mechanism and a Reason-in-Documents module to improve performance on complex reasoning tasks.

105Large reasoning modelsreinforcement learningHF ↗arXiv ↗
09

Learning to Reason under Off-Policy Guidance

Jianhao Yan, Yafu Li, Zican Hu +5 authors

LUFFY enhances zero-RL models with off-policy guidance, improving reasoning and generalization through balanced imitation and exploration.

88large reasoning modelsreinforcement learningHF ↗arXiv ↗
10

S*: Test Time Scaling for Code Generation

Dacheng Li, Shiyi Cao, Chengkun Cao +6 authors

A hybrid test-time scaling framework improves code generation coverage and accuracy across various models and domains.

63hybrid test-time scaling frameworkparallel scalingHF ↗arXiv ↗

上升最快

近 6 个月
1
35 篇论文
2
llmNEW
34 篇论文
3
29 篇论文
4
26 篇论文
5
ditNEW
12 篇论文
6
12 篇论文
7
12 篇论文
8
12 篇论文
9
11 篇论文
10
11 篇论文
11
11 篇论文
12
10 篇论文
13
10 篇论文
14
10 篇论文
15
10 篇论文
16
26 篇论文
17
74 篇论文
19
rlvr+200%
13 篇论文
20
12 篇论文

最热方向

按总量
1
3
167 篇论文
5
75 篇论文
6
74 篇论文
10
49 篇论文
11
39 篇论文
12
38 篇论文
13
14
15
35 篇论文
16
34 篇论文
17
33 篇论文
18
29 篇论文
19
29 篇论文
20
29 篇论文
21
28 篇论文
23
27 篇论文
24
27 篇论文
25
26 篇论文
26
29
25 篇论文
30
24 篇论文
31
23 篇论文
32
23 篇论文
33
23 篇论文
34
22 篇论文
35
22 篇论文
36
20 篇论文
37
20 篇论文
38
20 篇论文
39
20 篇论文
40
20 篇论文
41
19 篇论文
42
19 篇论文
43
19 篇论文
46
18 篇论文
48
51
53
55
16 篇论文
56
16 篇论文
58
59
16 篇论文
60

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号