TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Nov 20 – Nov 26, 2023
本周最热253

GAIA: a benchmark for General AI Assistants

Grégoire Mialon, Clémentine Fourrier, Craig Swift +3 authors

GAIA benchmarks general AI assistants using real-world questions that challenge both reasoning and multi-modality handling, showcasing a significant gap between human and AI performance.

multi-modality handlingweb browsingtool-use proficiencyGPT-4HF ↗arXiv ↗

44 篇论文 · 按点赞排序

03

Orca 2: Teaching Small Language Models How to Reason

Arindam Mitra, Luciano Del Corro, Shweti Mahajan +12 authors

Orca 2 enhances smaller language models' reasoning abilities by teaching them diverse solution strategies, outperforming larger models on complex reasoning tasks.

77imitation learningreasoning techniquesHF ↗arXiv ↗
04

Make Pixels Dance: High-Dynamic Video Generation

Yan Zeng, Guoqiang Wei, Jiani Zheng +4 authors

PixelDance, a diffusion model-based approach, generates high-dynamic videos by incorporating image instructions for first and last frames alongside text instructions, surpassing current text-to-video methods in complexity and motion.

67diffusion modelsimage instructionsHF ↗arXiv ↗
07

Diffusion Model Alignment Using Direct Preference Optimization

Bram Wallace, Meihua Dang, Rafael Rafailov +7 authors

A method called Diffusion-DPO aligns text-to-image diffusion models to human preferences using direct optimization on comparison data, improving visual appeal and prompt alignment.

48Reinforcement Learning from Human Feedback (RLHF)human comparison dataHF ↗arXiv ↗
10

GPQA: A Graduate-Level Google-Proof Q&A Benchmark

David Rein, Betty Li Hou, Asa Cooper Stickland +5 authors

A dataset of extremely difficult multiple-choice questions challenges both experts and AI systems, facilitating the development of scalable oversight methods for AI-generated knowledge.

38GPQAmultiple-choice questionsHF ↗arXiv ↗
23

PhysGaussian: Physics-Integrated 3D Gaussians for Generative Dynamics

Tianyi Xie, Zeshun Zong, Yuxin Qiu +4 authors

PhysGaussian integrates Newtonian dynamics within 3D Gaussians for high-quality motion synthesis using a Material Point Method (MPM) that aligns with continuum mechanics principles, eliminating the need for traditional meshing techniques.

21PhysGaussianMaterial Point Method (MPM)HF ↗arXiv ↗
26

Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2

Hamish Ivison, Yizhong Wang, Valentina Pyatkin +8 authors

T\"ULU 2, an advanced suite of language models, achieves state-of-the-art performance through improved datasets, fine-tuning techniques, and direct preference optimization, surpassing even GPT-3.5-turbo-0301 on several benchmarks.

19instruction tuningT\"ULUHF ↗arXiv ↗
27

Visual In-Context Prompting

Feng Li, Qing Jiang, Hao Zhang +9 authors

A universal visual in-context prompting framework enhances zero-shot capabilities for both referring and generic vision tasks by using a versatile prompt encoder and reference image segments as context.

18in-context promptinglarge language models (LLMs)HF ↗arXiv ↗
28

PG-Video-LLaVA: Pixel Grounding Large Video-Language Models

Shehan Munasinghe, Rusiru Thushara, Muhammad Maaz +4 authors

Video-LLaVA integrates audio transcriptions and uses a novel grounding module to enhance pixel-level grounding in large multimodal video models, leading to improved performance in generative and question-answering tasks.

18Large Multimodal ModelsLMMHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号