TensorX

Trends · 研究趋势

数据来自 Hugging Face 论文的 AI 提取关键词,按月统计研究方向的增长与热度。

返回趋势

large language models 相关论文

319 篇论文 · 按点赞排序

166

Hermes 3 Technical Report

Ryan Teknium, Jeffrey Quesnelle, Chen Guang

Hermes 3, a neutrally-aligned instruct and tool use model with strong reasoning and creative capabilities, achieves top performance on public benchmarks.

60instruct-tuned modelslarge language modelsHF ↗arXiv ↗
168

Fino1: On the Transferability of Reasoning Enhanced LLMs to Finance

Lingfei Qian, Weipeng Zhou, Yan Wang +3 authors

A study evaluates 16 large language models on complex financial tasks, finding that domain-specific CoT fine-tuning and reinforcement learning improve performance and highlight the need for further research on long-context and multi-table reasoning.

59large language modelsfinancial reasoningHF ↗arXiv ↗
171

Demystifying Long Chain-of-Thought Reasoning in LLMs

Edward Yeo, Yuxuan Tong, Morry Niu +2 authors

Investigation into long chains-of-thought reasoning in large language models reveals the critical role of training compute, reward shaping, and verifiable reward signals in enabling and measuring this capability.

57large language modelslong chains-of-thoughtHF ↗arXiv ↗
172

More Agents Is All You Need

Junyou Li, Qin Zhang, Yangbin Yu +2 authors

A sampling-and-voting method enhances large language models' performance by increasing the number of agents, with effectiveness tied to task difficulty.

57large language modelsLLMSHF ↗arXiv ↗
173

LLM360: Towards Fully Transparent Open-Source LLMs

Zhengzhong Liu, Aurick Qiao, Willie Neiswanger +25 authors

LLM360 initiative promotes full transparency and reproducibility in LLM training by open-sourcing training code, data, model checkpoints, and intermediate results.

57Large Language ModelsLLaMAHF ↗arXiv ↗
179

Inside-Out: Hidden Factual Knowledge in LLMs

Zorik Gekhman, Eyal Ben David, Hadas Orgad +5 authors

LLMs encode more internal factual knowledge than they express externally, with some knowledge so deeply hidden that it is never generated, despite repeated sampling.

56large language modelsLLMsHF ↗arXiv ↗
9 / 16

上升最快

近 6 个月
1
35 篇论文
2
llmNEW
34 篇论文
3
29 篇论文
4
26 篇论文
5
ditNEW
12 篇论文
6
12 篇论文
7
12 篇论文
8
12 篇论文
9
11 篇论文
10
11 篇论文
11
11 篇论文
12
10 篇论文
13
10 篇论文
14
10 篇论文
15
10 篇论文
16
26 篇论文
17
74 篇论文
19
rlvr+200%
13 篇论文
20
12 篇论文

最热方向

按总量
1
3
167 篇论文
5
75 篇论文
6
74 篇论文
10
49 篇论文
11
39 篇论文
12
38 篇论文
13
14
15
35 篇论文
16
34 篇论文
17
33 篇论文
18
29 篇论文
19
29 篇论文
20
29 篇论文
21
28 篇论文
23
27 篇论文
24
27 篇论文
25
26 篇论文
26
29
25 篇论文
30
24 篇论文
31
23 篇论文
32
23 篇论文
33
23 篇论文
34
22 篇论文
35
22 篇论文
36
20 篇论文
37
20 篇论文
38
20 篇论文
39
20 篇论文
40
20 篇论文
41
19 篇论文
42
19 篇论文
43
19 篇论文
46
18 篇论文
48
51
53
55
16 篇论文
56
16 篇论文
58
59
16 篇论文
60

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号