TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Jan 8 – Jan 14, 2024
本周最热161

Mixtral of Experts

Albert Q. Jiang, Alexandre Sablayrolles, Antoine Roux +23 authors

Mixtral 8x7B, a Sparse Mixture of Experts language model, achieves superior performance across benchmarks by using a selective architecture that leverages fewer active parameters.

Sparse Mixture of Experts (SMoE)feedforward blocksexpertsrouter networkHF ↗arXiv ↗

50 篇论文 · 按点赞排序

03

TrustLLM: Trustworthiness in Large Language Models

Lichao Sun, Yue Huang, Haoran Wang +64 authors

This study assesses the trustworthiness of large language models across various dimensions, including truthfulness, safety, fairness, robustness, privacy, and machine ethics, finding a positive correlation with utility and highlighting differences between proprietary and open-source models.

69TrustLLMlarge language modelsHF ↗arXiv ↗
08

MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation

Weimin Wang, Jiawei Liu, Zhijie Lin +9 authors

MagicVideo-V2 generates high-fidelity and smooth videos from text using an integrated pipeline that includes text-to-image, video motion generation, and frame interpolation modules, outperforming existing systems in user evaluations.

49text-to-imagevideo motion generatorHF ↗arXiv ↗
11

Transformers are Multi-State RNNs

Matanel Oren, Michael Hassid, Yossi Adi +1 authors

Decoder-only transformers can be conceptualized as finite multi-state RNNs, and a new cache compression technique, TOVA, significantly reduces their computational cost while maintaining high performance.

39transformersrecurrent neural networksHF ↗arXiv ↗
13

Denoising Vision Transformers

Jiawei Yang, Katie Z Luo, Jiefeng Li +3 authors

A novel noise model is proposed to eliminate grid-like artifacts in Vision Transformers by decomposing outputs into noise-free and artifact-related components, enhancing performance in downstream tasks.

32Vision TransformersViTsHF ↗arXiv ↗
20

URHand: Universal Relightable Hands

Zhaoxi Chen, Gyeongsik Moon, Kaiwen Guo +20 authors

URHand is a universal relightable hand model that generalizes across different identities, viewpoints, and illuminations, using a spatially varying linear lighting model and joint learning of a physically based model for high-quality, real-time rendering.

25neural relightinglight stageHF ↗arXiv ↗
26

TOFU: A Task of Fictitious Unlearning for LLMs

Pratyush Maini, Zhili Feng, Avi Schwarzschild +2 authors

TOFU, a benchmark for evaluating unlearning in large language models, uses synthetic data to measure how effectively unlearning methods can remove specific information from the model.

20large language modelsunlearningHF ↗arXiv ↗
28

Jump Cut Smoothing for Talking Heads

Xiaojuan Wang, Taesung Park, Yang Zhou +2 authors

A framework for smoothing jump cuts in talking head videos using DensePose keypoints, face landmarks, and cross-modal attention for seamless transitions.

20DensePose keypointsface landmarksHF ↗arXiv ↗
29

Towards Conversational Diagnostic AI

Tao Tu, Anil Palepu, Mike Schaekermann +22 authors

An AI system named AMIE, designed with a large language model optimized for diagnostic dialogue, demonstrated better performance than primary care physicians in several clinical axes through a simulated environment and textual consultations with patient actors.

19Large Language Modelself-playHF ↗arXiv ↗
30

Distilling Vision-Language Models on Millions of Videos

Yue Zhao, Long Zhao, Xingyi Zhou +9 authors

A video-language model, fine-tuned from an image-language baseline with synthesized data, outperforms existing methods on various benchmarks by leveraging auto-generated captions for better textual supervision.

19vision-language modelsvideo-language modelsHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号