TensorX

Trends · 研究趋势

数据来自 Hugging Face 论文的 AI 提取关键词,按月统计研究方向的增长与热度。

返回趋势

large language models 相关论文

319 篇论文 · 按点赞排序

281

Learning From Mistakes Makes LLM Better Reasoner

Shengnan An, Zexiong Ma, Zeqi Lin +3 authors

LeMa, a learning-from-mistakes approach, enhances LLMs' mathematical reasoning by learning from inaccurate reasoning paths corrected by GPT-4, surpassing SOTA performance on math problems.

29Large language modelsLearning from MistakesHF ↗arXiv ↗
282

Imp: Highly Capable Large Multimodal Models for Mobile Devices

Zhenwei Shao, Zhou Yu, Jun Yu +5 authors

A systematic study of lightweight large multimodal models (LMMs) led to the development of the Imp family, which achieves superior performance compared to larger models and is capable of high-speed inference on mobile devices.

29large language modelslarge multimodal modelsHF ↗arXiv ↗
289

LIMA: Less Is More for Alignment

Chunting Zhou, Pengfei Liu, Puxin Xu +12 authors

A 65B parameter LLaMa language model trained with minimal instructional data matches or outperforms models with extensive human preference modeling in most cases, indicating that pretraining is predominantly responsible for knowledge acquisition.

27large language modelsunsupervised pretrainingHF ↗arXiv ↗
290

HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(ision), LLaVA-1.5, and Other Multi-modality Models

Fuxiao Liu, Tianrui Guan, Zongxia Li +4 authors

HallusionBench is a benchmark that highlights language hallucination and visual illusion issues in vision-language models (VLMs), showcasing the limitations of current state-of-the-art models like GPT-4V and LLaVA-1.5.

27Large language modelsvision modelsHF ↗arXiv ↗
291

CapsFusion: Rethinking Image-Text Data at Scale

Qiying Yu, Quan Sun, Xiaosong Zhang +4 authors

CapsFusion is an advanced framework that improves multimodal pretraining data by combining web-based image-text pairs and synthetic captions, leading to enhanced model performance, sample efficiency, and scalability.

27multimodal modelszero-shotHF ↗arXiv ↗
292

PolyLM: An Open Source Polyglot Large Language Model

Xiangpeng Wei, Haoran Wei, Huan Lin +15 authors

PolyLM, a multilingual LLM trained on 640 billion tokens, enhances multilingual capabilities through bilingual data and curriculum learning, outperforming other models on multilingual tasks while maintaining English performance.

27large language modelsmultilingual LLMHF ↗arXiv ↗
298

Platypus: Quick, Cheap, and Powerful Refinement of LLMs

Ariel N. Lee, Cole J. Hunter, Nataniel Ruiz

A fine-tuned and merged family of large language models named Platypus, using LoRA modules and a curated dataset, achieves top performance on the Open LLM Leaderboard with reduced data and compute.

25Large Language ModelsLLMsHF ↗arXiv ↗
299

How is ChatGPT's behavior changing over time?

Lingjiao Chen, Matei Zaharia, James Zou

The performance and behavior of GPT-3.5 and GPT-4 fluctuated significantly between March and June 2023 across various tasks, emphasizing the necessity for ongoing LLM quality monitoring.

25large language modelsLLMHF ↗arXiv ↗
15 / 16

上升最快

近 6 个月
1
35 篇论文
2
llmNEW
34 篇论文
3
29 篇论文
4
26 篇论文
5
ditNEW
12 篇论文
6
12 篇论文
7
12 篇论文
8
12 篇论文
9
11 篇论文
10
11 篇论文
11
11 篇论文
12
10 篇论文
13
10 篇论文
14
10 篇论文
15
10 篇论文
16
26 篇论文
17
74 篇论文
19
rlvr+200%
13 篇论文
20
12 篇论文

最热方向

按总量
1
3
167 篇论文
5
75 篇论文
6
74 篇论文
10
49 篇论文
11
39 篇论文
12
38 篇论文
13
14
15
35 篇论文
16
34 篇论文
17
33 篇论文
18
29 篇论文
19
29 篇论文
20
29 篇论文
21
28 篇论文
23
27 篇论文
24
27 篇论文
25
26 篇论文
26
29
25 篇论文
30
24 篇论文
31
23 篇论文
32
23 篇论文
33
23 篇论文
34
22 篇论文
35
22 篇论文
36
20 篇论文
37
20 篇论文
38
20 篇论文
39
20 篇论文
40
20 篇论文
41
19 篇论文
42
19 篇论文
43
19 篇论文
46
18 篇论文
48
51
53
55
16 篇论文
56
16 篇论文
58
59
16 篇论文
60

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号