TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Oct 2 – Oct 8, 2023
本周最热79

Kandinsky: an Improved Text-to-Image Synthesis with Image Prior and Latent Diffusion

Anton Razzhigaev, Arseniy Shakhmatov, Anastasia Maltseva +7 authors

Kandinsky1, a latent diffusion architecture, achieves high-quality text-to-image generation by integrating image prior models and modified MoVQ autoencoders, outperforming other open-source models.

diffusion-based modelspixel-levellatent-levelimage prior modelsHF ↗arXiv ↗

21 篇论文 · 按点赞排序

09

Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Mihir Prabhudesai, Anirudh Goyal, Deepak Pathak +1 authors

AlignProp refines text-to-image diffusion models using backpropagation through the denoising process, leveraging low-rank adapters and gradient checkpointing to optimize for various objectives with higher efficiency.

22text-to-image diffusion modelsreinforcement learningHF ↗arXiv ↗
13

Conditional Diffusion Distillation

Kangfu Mei, Mauricio Delbracio, Hossein Talebi +3 authors

A novel single-stage distillation method for generative diffusion models reduces sampling time while maintaining performance across tasks like super-resolution and image editing.

19generative diffusion modelstext-to-image generationHF ↗arXiv ↗
14

Large Language Models as Analogical Reasoners

Michihiro Yasunaga, Xinyun Chen, Yujia Li +5 authors

Analogical Prompting enhances language model performance in reasoning tasks by self-generating relevant exemplars without labeled data.

16Analogical Promptinganalogical reasoningHF ↗arXiv ↗
17

SmartPlay : A Benchmark for LLMs as Intelligent Agents

Yue Wu, Xuan Tang, Tom M. Mitchell +1 authors

SmartPlay is a benchmark and evaluation methodology for large language models as agents, featuring 6 games that test various capabilities like reasoning, planning, spatial awareness, and learning.

13large language modelsintelligent agentsHF ↗arXiv ↗
18

A Long Way to Go: Investigating Length Correlations in RLHF

Prasann Singhal, Tanya Goyal, Jiacheng Xu +1 authors

Optimizing for response length significantly contributes to the improvements observed in Reinforcement Learning from Human Feedback (RLHF) when aligning large language models for helpfulness tasks.

10Reinforcement Learning from Human Feedback (RLHF)reward modelsHF ↗arXiv ↗
19

Drag View: Generalizable Novel View Synthesis with Unposed Imagery

Zhiwen Fan, Panwang Pan, Peihao Wang +6 authors

DragView is an interactive framework that generates novel views of unseen scenes using a single source image and sparse unposed multi-view images, utilizing an epipolar attention mechanism and transformer-based ray decoding without estimating camera poses.

8interactive frameworkDragViewHF ↗arXiv ↗

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号