TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Aug 5 – Aug 11, 2024
本周最热175

Transformer Explainer: Interactive Learning of Text-Generative Models

Aeree Cho, Grace C. Kim, Alexander Karpekov +5 authors

Transformer Explainer is an interactive visualization tool that allows non-experts to understand the inner workings of the GPT-2 model through real-time experimentation and visualization in a web browser.

TransformersGPT-2interactive visualizationmodel overviewHF ↗arXiv ↗

50 篇论文 · 按点赞排序

02

MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Yuan Yao, Tianyu Yu, Ao Zhang +20 authors

MiniCPM-V presents a series of efficient Multimodal Large Language Models optimized for end-side deployment, offering high performance and practical usability compared to larger models.

95Multimodal Large Language ModelsMLLMsHF ↗arXiv ↗
05

LLaVA-OneVision: Easy Visual Task Transfer

Bo Li, Yuanhan Zhang, Dong Guo +7 authors

LLaVA-OneVision is a unified multimodal model that advances performance across single-image, multi-image, and video scenarios with strong transfer learning capabilities.

61multimodal modelsLMMsHF ↗arXiv ↗
09

Language Model Can Listen While Speaking

Ziyang Ma, Yakun Song, Chenpeng Du +5 authors

A novel listening-while-speaking language model (LSLM) enhances real-time, full-duplex speech interaction by integrating speech generation and real-time audio input with an optimal fusion strategy.

40full duplex modelinginteractive speech language modelsHF ↗arXiv ↗
11

EXAONE 3.0 7.8B Instruction Tuned Language Model

LG AI Research, Soyoung An, Kyunghoon Bae +35 authors

EXAONE 3.0, a 7.8B instruction-tuned language model from LG AI Research, demonstrates strong performance and instruction-following capability, particularly in Korean and across general complex reasoning tasks.

37instruction-tuned language modelLarge Language Models (LLMs)HF ↗arXiv ↗
16

Achieving Human Level Competitive Robot Table Tennis

David B. D'Ambrosio, Saminda Abeyruwan, Laura Graesser +24 authors

A robot agent achieves amateur human-level performance in competitive table tennis through a hierarchical policy architecture and real-time adaptation to unseen opponents.

28hierarchical and modular policy architecturelow level controllersHF ↗arXiv ↗
17

Self-Taught Evaluators

Tianlu Wang, Ilia Kulikov, Olga Golovneva +7 authors

A self-improving evaluator trained with synthetic data outperforms traditional human-annotated evaluators and matches top-performing reward models on RewardBench.

28LLM-as-a-Judgereasoning tracesHF ↗arXiv ↗
18

POA: Pre-training Once for Models of All Sizes

Yingying Zhang, Xin Guo, Jiangwei Lao +7 authors

POA framework facilitates the pre-training of various-sized models through an innovative elastic student branch within self-distillation, achieving state-of-the-art performance across multiple backbones.

27self-supervised pre-trainingtri-branch frameworkHF ↗arXiv ↗
22

Trans-Tokenization and Cross-lingual Vocabulary Transfers: Language Adaptation of LLMs for Low-Resource NLP

François Remy, Pieter Delobelle, Hayastan Avetisyan +3 authors

A novel trans-tokenization strategy is introduced to transfer vocabulary from high-resource languages to low-resource languages, enabling competitive performance in various tasks and facilitating zero-shot machine translation for languages like Tatar.

21monolingual language modelslow-resource languagesHF ↗arXiv ↗
26

Unleashing the Power of Data Tsunami: A Comprehensive Survey on Data Assessment and Selection for Instruction Tuning of Language Models

Yulei Qin, Yuncheng Yang, Pengcheng Guo +7 authors

A review of existing data assessment and selection methods for instruction tuning of large language models categorizes them into quality-based, diversity-based, and importance-based approaches, discusses their limitations, and outlines future research directions.

18large language modelsinstruction tuningHF ↗arXiv ↗
29

Diffusion Models as Data Mining Tools

Ioannis Siglidis, Aleksander Holynski, Alexei A. Efros +2 authors

The approach uses fine-tuned conditional diffusion models to define a typicality measure for visual data elements across various labels and datasets, improving scalability and versatility in visual data mining.

15generative modelsimage synthesisHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号