TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Sep 11 – Sep 17, 2023
本周最热92

Textbooks Are All You Need II: phi-1.5 technical report

Yuanzhi Li, Sébastien Bubeck, Ronen Eldan +3 authors

A new 1.3 billion parameter Transformer-based language model, phi-1.5, demonstrates comparable performance to much larger models on common sense reasoning and complex tasks despite the absence of web data.

Transformer-based language modelsTinyStoriesphi-1Large Language Models (LLMs)HF ↗arXiv ↗

36 篇论文 · 按点赞排序

02

NExT-GPT: Any-to-Any Multimodal LLM

Shengqiong Wu, Hao Fei, Leigang Qu +2 authors

NExT-GPT, an any-to-any Multimodal Large Language Model, combines LLMs with multimodal adaptors and diffusion decoders to generate content across various modalities, enhanced by modality-switching instruction tuning and a curated dataset.

79Multimodal Large Language ModelsMM-LLMsHF ↗arXiv ↗
04

Large-Scale Automatic Audiobook Creation

Brendan Walsh, Mark Hamilton, Greg Newby +8 authors

A system leverages neural text-to-speech to automatically generate high-quality, customizable audiobooks from e-books, significantly expanding access to literature.

55neural text-to-speechHF ↗arXiv ↗
05

Generative Image Dynamics

Zhengqi Li, Richard Tucker, Noah Snavely +1 authors

A frequency-coordinated diffusion sampling process is used to predict long-term motion representations for still images, enabling dynamic video creation and interactive scene manipulation.

54frequency-coordinated diffusion sampling processneural stochastic motion textureHF ↗arXiv ↗
07

Agents: An Open-source Framework for Autonomous Language Agents

Wangchunshu Zhou, Yuchen Eleanor Jiang, Long Li +14 authors

Agents is an open-source library that facilitates the creation, customization, and deployment of autonomous language agents with features like planning, memory, and tool usage, aiming to make advances in large language models accessible to a broader audience.

43large language modelsautonomous language agentsHF ↗arXiv ↗
10

AudioSR: Versatile Audio Super-resolution at Scale

Haohe Liu, Ke Chen, Qiao Tian +2 authors

A diffusion-based generative model, AudioSR, achieves robust audio super-resolution across various audio types and bandwidths, enhancing generation quality for different audio models.

28diffusion-based generative modelaudio super-resolutionHF ↗arXiv ↗
13

MADLAD-400: A Multilingual And Document-Level Large Audited Dataset

Sneha Kudugunta, Isaac Caswell, Biao Zhang +8 authors

A multilingual translation model trained on a large dataset of 250 billion tokens across over 450 languages outperforms larger models, with competitive results in various domains and effective few-shot translation performance.

26multilingual machine translationfew-shot translationHF ↗arXiv ↗
14

Large Language Models for Compiler Optimization

Chris Cummins, Volker Seeker, Dejan Grubisic +8 authors

A transformer model trained to optimize LLVM assembly code size significantly improves compiler performance with auxiliary learning tasks, achieving better results than baselines and demonstrating strong code reasoning abilities.

25Large Language Modelstransformer modelHF ↗arXiv ↗
17

Neurons in Large Language Models: Dead, N-gram, Positional

Elena Voita, Javier Ferrando, Christoforos Nalmpantis

The analysis of the OPT family of large language models reveals that early network layers are sparse, with many neurons never activating, and some neurons functioning as token detectors that also remove information about triggering tokens.

18OPT familyFFN neuronsHF ↗arXiv ↗
18

Dynamic NeRFs for Soccer Scenes

Sacha Lewin, Maxime Vandegar, Thomas Hoyoux +2 authors

Dynamic NeRFs show promise for photorealistic novel view synthesis of soccer scenes, though they do not yet meet the quality standards required by the broadcast industry.

17neural radiance fieldsNeRFsHF ↗arXiv ↗
22

Statistical Rejection Sampling Improves Preference Optimization

Tianqi Liu, Yao Zhao, Rishabh Joshi +4 authors

A novel approach called Statistical Rejection Sampling Optimization (RSO) enhances preference data sourcing for aligning language models with human preferences, outperforming previous methods like SLiC and DPO.

15Reinforcement Learning from Human Feedback (RLHF)Proximal Policy Optimization (PPO)HF ↗arXiv ↗
23

Uncovering mesa-optimization algorithms in Transformers

Johannes von Oswald, Eyvind Niklasson, Maximilian Schlegel +9 authors

Transformers' superior performance might be due to mesa-optimization, an internal learning objective and its solution during the forward pass, which can be repurposed for few-shot tasks and potentially enhances language models' in-context learning.

15Transformersmesa-optimizationHF ↗arXiv ↗
27

Towards Practical Capture of High-Fidelity Relightable Avatars

Haotian Yang, Mingwu Zheng, Wanquan Feng +5 authors

A tracking-free framework captures and reconstructs high-fidelity 3D avatars with realistic relighting and real-time animation using a novel network architecture optimized for capturing dynamic lighting conditions.

10Tracking-free avatar captureLight StageHF ↗arXiv ↗
30

Dynamic Mesh-Aware Radiance Fields

Yi-Ling Qiao, Alexander Gao, Yiran Xu +3 authors

Rendering and simulating mesh assets within NeRF volumes integrates mesh and NeRF light transport in an efficient, GPU-accelerated system for enhanced visual realism.

7Neural Radience FieldsNeRFHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号