TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Jul 17 – Jul 23, 2023
本周最热252

Llama 2: Open Foundation and Fine-Tuned Chat Models

Hugo Touvron, Louis Martin, Kevin Stone +65 authors

Llama 2, a series of pretrained and fine-tuned large language models, achieves superior performance in dialogue tasks compared to open-source alternatives and offers safety enhancements.

pretrained language modelsfine-tuningdialogue use caseshuman evaluationsHF ↗arXiv ↗

42 篇论文 · 按点赞排序

04

Challenges and Applications of Large Language Models

Jean Kaddour, Joshua Harris, Maximilian Mozes +3 authors

Large Language Models (LLMs) went from non-existent to ubiquitous in the machine learning discourse within a few years. Due to the fast pace of the field, it is difficult to identify the remaining challenges and already fruitful application areas. In this paper, we aim to establish a systematic set of open problems and application successes so that ML researchers can comprehend the field's current state more quickly and become productive.

51Large Language Models (LLMs)discourseHF ↗arXiv ↗
08

Brain2Music: Reconstructing Music from Human Brain Activity

Timo I. Denk, Yu Takagi, Takuya Matsuyama +4 authors

A method reconstructs music from fMRI data using the MusicLM model, matching semantic properties of original stimuli and identifying brain regions involved in processing music descriptions.

42functional magnetic resonance imagingfMRIHF ↗arXiv ↗
09

Copy Is All You Need

Tian Lan, Deng Cai, Yan Wang +2 authors

Text generation is improved by copying text segments from a pre-existing collection, resulting in better quality and efficiency compared to traditional methods.

36contextualized representationsvector search toolkitsHF ↗arXiv ↗
13

How is ChatGPT's behavior changing over time?

Lingjiao Chen, Matei Zaharia, James Zou

The performance and behavior of GPT-3.5 and GPT-4 fluctuated significantly between March and June 2023 across various tasks, emphasizing the necessity for ongoing LLM quality monitoring.

26large language modelsLLMHF ↗arXiv ↗
17

Diffusion Models Beat GANs on Image Classification

Soumik Mukhopadhyay, Matthew Gwilliam, Vatsal Agarwal +5 authors

Diffusion models trained on image generation tasks produce useful feature embeddings that perform well for classification tasks and outperform generative-discriminative models like BigBiGAN.

20diffusion modelsgenerative tasksHF ↗arXiv ↗
18

CoTracker: It is Better to Track Together

Nikita Karaev, Ignacio Rocco, Benjamin Graham +3 authors

CoTracker, a deep learning architecture combining transformer networks and specialized attention layers, jointly tracks multiple video points across frames, offering superior performance compared to existing methods.

19optical flowtrackingHF ↗arXiv ↗
22

Towards A Unified Agent with Foundation Models

Norman Di Palo, Arunkumar Byravan, Leonard Hasenclever +3 authors

A framework using language as a reasoning tool in RL agents improves exploration efficiency and data reuse, and enhances the acquisition of skills for novel tasks.

14language modelsvision language modelsHF ↗arXiv ↗
24

Improving Multimodal Datasets with Image Captioning

Thao Nguyen, Samir Yitzhak Gadre, Gabriel Ilharco +2 authors

Generated captions improve the utility of web-scraped image-text datasets by reducing noise without compromising diversity, outperforming existing filtering methods across multiple benchmarks and tasks.

12vision-language modelsCLIPHF ↗arXiv ↗
26

Does Circuit Analysis Interpretability Scale? Evidence from Multiple Choice Capabilities in Chinchilla

Tom Lieberum, Matthew Rahtz, János Kramár +3 authors

Circuit analysis applied to a 70B Chinchilla model provides insights into the mechanisms of multiple-choice question answering, demonstrating that logit attribution, attention pattern visualization, and activation patching scale and allow identification of output nodes like attention heads and MLPs.

12logit attributionattention pattern visualizationHF ↗arXiv ↗
27

Planting a SEED of Vision in Large Language Model

Yuying Ge, Yixiao Ge, Ziyun Zeng +2 authors

SEED, an image tokenizer, enables Large Language Models to perform image-to-text and text-to-image generation by incorporating independent 1D image tokens optimized for semantic alignment and reconstruction.

12SEEDimage tokenizerHF ↗arXiv ↗
28

PASTA: Pretrained Action-State Transformer Agents

Raphael Boige, Yannis Flet-Berliac, Arthur Flajolet +2 authors

Pretrained Action-State Transformer Agents (PASTA) are investigated for a range of reinforcement learning tasks using unified methodologies and parameter-efficient fine-tuning.

11self-supervised learningtransformer modelsHF ↗arXiv ↗
30

INVE: Interactive Neural Video Editing

Jiahui Huang, Leonid Sigal, Kwang Moo Yi +2 authors

Interactive Neural Video Editing (INVE) improves real-time video editing by leveraging hash-grids encoding and bi-directional functions, enabling faster processing and a wider range of edit operations compared to Layered Neural Atlas (LNA).

11Interactive Neural Video EditingINVEHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号