TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

50 篇论文 · 按点赞排序

32

Benchmarking Neural Network Training Algorithms

George E. Dahl, Frank Schneider, Zachary Nado +22 authors

A new benchmark, AlgoPerf: Training Algorithms, addresses challenges in evaluating training algorithms by providing a competitive, time-to-result benchmark across workloads and optimizers.

22update rulestuning protocolsHF ↗arXiv ↗
34

Demystifying GPT Self-Repair for Code Generation

Theo X. Olausson, Jeevana Priya Inala, Chenglong Wang +2 authors

GPT-4 demonstrates superior self-repair capabilities on APPS dataset compared to GPT-3.5, with performance boosted when feedback is provided by GPT-4 or human programmers.

21Large Language ModelsLLMsHF ↗arXiv ↗
40

h2oGPT: Democratizing Large Language Models

Arno Candel, Jon McKinney, Philipp Singer +12 authors

h2oGPT provides open-source, fine-tuned LLMs based on Generative Pretrained Transformers with 100% private document search capabilities.

19Generative Pretrained TransformersLLMsHF ↗arXiv ↗
42

Augmenting Language Models with Long-Term Memory

Weizhi Wang, Li Dong, Hao Cheng +4 authors

A framework called LongMem enables large language models to utilize long-term memory, overcoming input length limitations and improving performance on long-context tasks.

19language modelslong-term memoryHF ↗arXiv ↗
46

Scaling MLPs: A Tale of Inductive Bias

Gregor Bachmann, Sotiris Anagnostidis, Thomas Hofmann

MLPs achieve competitive performance on vision tasks with large-scale pre-training, challenging the narrative that inductive bias is necessary for high accuracy.

17multi-layer perceptron (MLP)inductive biasHF ↗arXiv ↗
47

HomeRobot: Open-Vocabulary Mobile Manipulation

Sriram Yenamandra, Arun Ramachandran, Karmesh Yadav +15 authors

The HomeRobot OVMM benchmark evaluates robots' ability to manipulate unseen objects in new environments using a combination of simulated and real-world testing.

17Open-Vocabulary Mobile ManipulationpercpetionHF ↗arXiv ↗
48

DreamHuman: Animatable 3D Avatars from Text

Nikos Kolotouros, Thiemo Alldieck, Andrei Zanfir +3 authors

DreamHuman generates realistic animatable 3D human avatars using text by integrating text-to-image synthesis, neural radiance fields, and statistical human body models.

17text-to-3D methodsneural radiance fieldsHF ↗arXiv ↗
49

Scalable 3D Captioning with Pretrained Models

Tiange Luo, Chris Rockwell, Honglak Lee +1 authors

Cap3D generates high-quality descriptive text for 3D objects using pretrained models and datasets, surpassing human performance in quality, cost, and speed.

17image captioningimage-text alignmentHF ↗arXiv ↗
2 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号