TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

50 篇论文 · 按点赞排序

32

Video Diffusion Alignment via Reward Gradients

Mihir Prabhudesai, Russell Mendonca, Zheyang Qin +2 authors

Utilizing pre-trained reward models to adapt video diffusion models with gradient-based feedback enhances efficiency and performance compared to gradient-free methods.

49video diffusion modelspre-trained reward modelsHF ↗arXiv ↗
33

TabReD: A Benchmark of Tabular Machine Learning in-the-Wild

Ivan Rubachev, Nikolay Kartashev, Yury Gorishniy +1 authors

TabReD is a new collection of industry-grade tabular datasets that address temporal dynamics and feature engineering, revealing that simpler models outperform more complex deep learning architectures in these settings.

49TabReDMLP-like architecturesHF ↗arXiv ↗
38

Very Large-Scale Multi-Agent Simulation in AgentScope

Xuchen Pan, Dawei Gao, Yuexiang Xie +5 authors

Enhancements to the AgentScope platform improve scalability, efficiency, and ease of use for large-scale multi-agent simulations through distributed mechanisms, flexible environments, and user-friendly tools.

46actor-based distributed mechanismmulti-agent platformHF ↗arXiv ↗
41

EVLM: An Efficient Vision-Language Model for Visual Understanding

Kaibing Chen, Dong Shen, Hanwen Zhong +14 authors

A multi-modal language model using cross-attention, hierarchical ViT features, and Mixture of Experts mechanism achieves competitive performance in image and video captioning tasks with reduced computational costs.

44cross-attentionhierarchical ViT featuresHF ↗arXiv ↗
42

MindSearch: Mimicking Human Minds Elicits Deep AI Searcher

Zehui Chen, Kuikun Liu, Qiuchen Wang +4 authors

MindSearch, an LLM-based multi-agent framework, improves web information seeking and integration through parallel processing and hierarchical retrieval, achieving better performance than existing solutions.

43Large Language ModelsLLMsHF ↗arXiv ↗
43

KAN or MLP: A Fairer Comparison

Runpeng Yu, Weihao Yu, Xinchao Wang

A comprehensive comparison of KAN and MLP models across diverse tasks reveals that MLP generally outperforms KAN except in symbolic formula representation where KAN's B-spline activation function provides advantage, and KAN exhibits more severe forgetting issues in class-incremental continual learning.

43KANMLPHF ↗arXiv ↗
48

SHIC: Shape-Image Correspondences with no Keypoint Supervision

Aleksandar Shtedritski, Christian Rupprecht, Andrea Vedaldi

SHIC leverages foundation computer vision models to learn canonical surface maps without manual supervision, achieving superior results by simulating the annotation process using image-to-image correspondences and enhanced template views.

41DensePosekeypoint detectionHF ↗arXiv ↗
49

VILA^2: VILA Augmented VILA

Yunhao Fang, Ligeng Zhu, Yao Lu +6 authors

A novel data augmentation approach iteratively improves visual language model data quality and performance using self-augmentation and specialist-augmentation, leading to state-of-the-art results on MMMU tasks.

41visual language modelslarge language modelsHF ↗arXiv ↗
2 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号