TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

May 22 – May 28, 2023
本周最热64

QLoRA: Efficient Finetuning of Quantized LLMs

Tim Dettmers, Artidoro Pagnoni, Ari Holtzman +1 authors

QLoRA enables efficient finetuning of large language models using 4-bit quantization and Low Rank Adapters, achieving high performance with reduced memory usage.

QLoRALow Rank AdaptersLoRA4-bit quantizedHF ↗arXiv ↗

50 篇论文 · 按点赞排序

02

LIMA: Less Is More for Alignment

Chunting Zhou, Pengfei Liu, Puxin Xu +12 authors

A 65B parameter LLaMa language model trained with minimal instructional data matches or outperforms models with extensive human preference modeling in most cases, indicating that pretraining is predominantly responsible for knowledge acquisition.

27large language modelsunsupervised pretrainingHF ↗arXiv ↗
03

RWKV: Reinventing RNNs for the Transformer Era

Bo Peng, Eric Alcaide, Quentin Anthony +27 authors

A new architecture, RWKV, combines the parallelizable training of Transformers with the efficient inference of RNNs, achieving linear scaling and matching performance.

21Transformersrecurrent neural networks (RNNs)HF ↗arXiv ↗
04

Voyager: An Open-Ended Embodied Agent with Large Language Models

Guanzhi Wang, Yuqi Xie, Yunfan Jiang +5 authors

Voyager is an LLM-powered agent in Minecraft that autonomously explores, learns, and discovers skills through a curriculum, skill library, and prompting mechanism, demonstrating superior lifelong learning and task-solving capabilities.

13LLM-poweredembodied lifelong learning agentHF ↗arXiv ↗
09

The False Promise of Imitating Proprietary LLMs

Arnav Gudibande, Eric Wallace, Charlie Snell +5 authors

Finetuning open-source language models on outputs from proprietary models like ChatGPT shows improved instruction-following but fails to close the performance gap on unsupported tasks.

6finetuneLMsHF ↗arXiv ↗
10

Is GPT-4 a Good Data Analyst?

Liying Cheng, Xingxuan Li, Lidong Bing

GPT-4 demonstrates comparable performance to human data analysts in end-to-end data analysis tasks across various domains.

6HF ↗arXiv ↗
13

On Architectural Compression of Text-to-Image Diffusion Models

Bo-Kyeong Kim, Hyoung-Kyu Song, Thibault Castells +1 authors

Classical architectural compression and knowledge distillation reduce the size and computational cost of Stable Diffusion models while maintaining competitive text-to-image synthesis performance.

5text-to-image (T2I) generationStable Diffusion models (SDMs)HF ↗arXiv ↗
15

Textually Pretrained Speech Language Models

Michael Hassid, Tal Remez, Tu Anh Nguyen +9 authors

TWIST improves SpeechLMs by leveraging pretrained textual language models, demonstrating superior performance across various evaluations and introducing new benchmarks for spoken data.

5SpeechLMswarm-startHF ↗arXiv ↗
18

Any-to-Any Generation via Composable Diffusion

Zineng Tang, Ziyi Yang, Chenguang Zhu +2 authors

CoDi is a generative model that can produce and condition on any combination of modalities by aligning them through a shared multimodal space in the diffusion process.

5Composable Diffusiongenerative modelHF ↗arXiv ↗
21

How Language Model Hallucinations Can Snowball

Muru Zhang, Ofir Press, William Merrill +2 authors

Language models frequently generate incorrect answers and offer false explanations, exacerbating errors due to a phenomenon called hallucination snowballing.

4hallucinationsknowledge gapsHF ↗arXiv ↗
26

Efficient Neural Music Generation

Max W. Y. Lam, Qiao Tian, Tang Li +10 authors

MeLoDy is an LM-guided diffusion model that generates high-quality music with reduced computational cost and improved sampling speed compared to MusicLM.

2MusicLMsemantic modelingHF ↗arXiv ↗
27

Manifold Diffusion Fields

Ahmed A. Elhag, Joshua M. Susskind, Miguel Angel Bautista

Manifold Diffusion Fields (MDF) learns generative models of continuous functions on Riemannian manifolds using eigen-functions of the Laplace-Beltrami Operator, achieving better diversity and fidelity.

2Manifold Diffusion FieldsMDFHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号