TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Dec 4 – Dec 10, 2023

50 篇论文 · 按点赞排序

31

Nash Learning from Human Feedback

Rémi Munos, Michal Valko, Daniele Calandriello +14 authors

NLHF uses a preference model and mirror descent to fine-tune LLMs for text summarization, advancing alignment with human preferences.

18reinforcement learning from human feedback (RLHF)large language models (LLMs)HF ↗arXiv ↗
32

MoMask: Generative Masked Modeling of 3D Human Motions

Chuan Guo, Yuxuan Mu, Muhammad Gohar Javed +2 authors

MoMask, a masked modeling framework using hierarchical quantization and bidirectional transformers, excels in text-to-motion generation and related tasks with high fidelity and competitive performance.

17masked modelinghierarchical quantizationHF ↗arXiv ↗
35

HiFi Tuner: High-Fidelity Subject-Driven Fine-Tuning for Diffusion Models

Zhonghao Wang, Wei Wei, Yang Zhao +4 authors

HiFi Tuner enhances object appearance conservation in personalized image generation through mask guidance, parameter regularization, and step-wise subject representations, achieving state-of-the-art results on the DreamBooth dataset.

16text-to-image diffusion modelsparameter-efficient fine-tuningHF ↗arXiv ↗
38

Context Diffusion: In-Context Aware Image Generation

Ivona Najdenkoska, Animesh Sinha, Abhimanyu Dubey +3 authors

Context Diffusion enhances in-context image generation by separately encoding visual context and preserving query image structure, improving quality and fidelity across different scenarios.

15diffusion-based frameworkvisual contextHF ↗arXiv ↗
46

Instruction-tuning Aligns LLMs to the Human Brain

Khai Loong Aw, Syrielle Montariol, Badr AlKhamissi +2 authors

Instruction-tuning improves brain alignment in large language models but not behavioral alignment, with correlations found between brain alignment and model size and world knowledge performance.

14instruction-tuninglarge language modelsHF ↗arXiv ↗
47

Dolphins: Multimodal Language Model for Driving

Yingzi Ma, Yulong Cao, Jiachen Sun +2 authors

Dolphins, a vision-language model enhanced with Grounded Chain of Thought, provides human-like capabilities as a conversational driving assistant by processing multimodal inputs and tailoring to specific driving tasks.

14Vision-Language ModelGrounded Chain of ThoughtHF ↗arXiv ↗
2 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号