TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Jul 22 – Jul 28, 2024

50 篇论文 · 按点赞排序

31

BOND: Aligning LLMs with Best-of-N Distillation

Pier Giuseppe Sessa, Robert Dadashi, Léonard Hussenot +17 authors

BOND, a novel RLHF algorithm, distills the Best-of-N sampling strategy without significant computational overhead, enhancing the quality of generative language models.

19Reinforcement learning from human feedbackRLHFHF ↗arXiv ↗
37

SciCode: A Research Coding Benchmark Curated by Scientists

Minyang Tian, Luyu Gao, Shizhuo Dylan Zhang +27 authors

SciCode, a scientist-curated coding benchmark, evaluates language models' ability to solve scientific research problems across various fields, highlighting current capabilities and future challenges in AI-assisted science.

17language modelscoding benchmarkHF ↗arXiv ↗
40

Discrete Flow Matching

Itai Gat, Tal Remez, Neta Shaul +5 authors

Discrete Flow Matching is a novel generative model for discrete data that achieves high performance without an autoregressive approach, improving benchmarks like HumanEval and MBPP.

14Flow Matchingdiffusion modelsHF ↗arXiv ↗
48

Consent in Crisis: The Rapid Decline of the AI Data Commons

Shayne Longpre, Robert Mahari, Ariel Lee +46 authors

The audit of web domains used for training AI systems reveals a growing inconsistency and restriction of data use, posing a significant threat to the diversity and availability of open web data for AI.

13Terms of Servicerobots.txtHF ↗arXiv ↗
50

Fast Matrix Multiplications for Lookup Table-Quantized LLMs

Han Guo, William Brandon, Radostin Cholakov +3 authors

FLUTE, a lookup table engine for quantized language models, accelerates inference by minimizing bit manipulations and optimizing shared memory usage, achieving faster performance compared to existing kernels.

13large language modelsweight-only quantizationHF ↗arXiv ↗
2 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号