TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

May 15 – May 21, 2023
本周最热48

TinyStories: How Small Can Language Models Be and Still Speak Coherent English?

Ronen Eldan, Yuanzhi Li

TinyStories, a synthetic dataset of simplified stories, enables small or simple language models to produce fluent and coherent text with reasoning capabilities using an evaluation framework based on human-like grading.

TinyStorieslanguage modelssynthetic datasetshort storiesHF ↗arXiv ↗

50 篇论文 · 按点赞排序

04

SoundStorm: Efficient Parallel Audio Generation

Zalán Borsos, Matt Sharifi, Damien Vincent +3 authors

SoundStorm, a non-autoregressive audio generation model, delivers high-quality and consistent audio two orders of magnitude faster than autoregressive methods.

15SoundStormnon-autoregressiveHF ↗arXiv ↗
06

LDM3D: Latent Diffusion Model for 3D

Gabriela Ben Melech Stan, Diana Wofk, Scottie Fox +8 authors

Latent Diffusion Model for 3D (LDM3D) generates RGBD images from text prompts and is used to create immersive 360-degree-view experiences.

13Latent Diffusion Model for 3DLDM3DHF ↗arXiv ↗
09

MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers

Lili Yu, Dániel Simig, Colin Flaherty +3 authors

Megabyte, a multi-scale decoder architecture, enables efficient byte-level modeling of long sequences through sub-quadratic self-attention, larger feedforward layers, and improved parallelism, achieving performance competitive with subword models and state-of-the-art results.

10Megabytemulti-scale decoder architectureHF ↗arXiv ↗
10

PaLM 2 Technical Report

Rohan Anil, Andrew M. Dai, Orhan Firat +125 authors

PaLM 2, a Transformer-based language model, improves multilingual and reasoning capabilities with enhanced efficiency and performance across various tasks compared to its predecessor.

9Transformer-based modelmixture of objectivesHF ↗arXiv ↗
12

SLiC-HF: Sequence Likelihood Calibration with Human Feedback

Yao Zhao, Rishabh Joshi, Tianqi Liu +3 authors

Sequence Likelihood Calibration (SLiC) is shown to be an effective and simpler alternative to Reinforcement Learning from Human Feedback (RLHF) for learning from human preferences in language models.

7Reinforcement Learning from Human Feedback (RLHF)Sequence Likelihood Calibration (SLiC)HF ↗arXiv ↗
23

TextDiffuser: Diffusion Models as Text Painters

Jingye Chen, Yupan Huang, Tengchao Lv +3 authors

TextDiffuser, a two-stage model combining Transformer and diffusion models, generates visually appealing and coherent text images using text prompts and layouts, leveraging the MARIO-10M dataset for training and MARIO-Eval for evaluation.

4TextDiffuserTransformer modelHF ↗arXiv ↗
26

FitMe: Deep Photorealistic 3D Morphable Model Avatars

Alexandros Lattas, Stylianos Moschoglou, Stylianos Ploumpis +3 authors

FitMe is a facial reflectance model with a differentiable rendering pipeline that generates high-fidelity human avatars from single or multiple images, offering accurate reflectance and identity preservation.

4facial reflectance modeldifferentiable rendering optimization pipelineHF ↗arXiv ↗
27

Online Continual Learning Without the Storage Constraint

Ameya Prabhu, Zhipeng Cai, Puneet Dokania +3 authors

Research addresses online continual learning with focus on computational budgets, using a kNN classifier and pre-trained feature extractors to manage large datasets efficiently.

4kNN classifieruniversal pre-trained feature extractorsHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号