TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

606 篇论文 · 按点赞排序

571

VACE: All-in-One Video Creation and Editing

Zeyinzi Jiang, Zhen Han, Chaojie Mao +3 authors

VACE, an all-in-one framework for video creation and editing, integrates multiple tasks within a unified model using a Video Condition Unit and Context Adapter for flexible and consistent video synthesis.

58diffusion transformervideo synthesisHF ↗arXiv ↗
575

Demystifying Long Chain-of-Thought Reasoning in LLMs

Edward Yeo, Yuxuan Tong, Morry Niu +2 authors

Investigation into long chains-of-thought reasoning in large language models reveals the critical role of training compute, reward shaping, and verifiable reward signals in enabling and measuring this capability.

57large language modelslong chains-of-thoughtHF ↗arXiv ↗
579

OpenThoughts: Data Recipes for Reasoning Models

Etash Guha, Ryan Marten, Sedrick Keh +47 authors

The OpenThoughts project created open-source datasets leading to reasoning models that match or exceed state-of-the-art benchmarks in math, code, and science.

56reasoning modelsOpenThoughts projectHF ↗arXiv ↗
582

Inside-Out: Hidden Factual Knowledge in LLMs

Zorik Gekhman, Eyal Ben David, Hadas Orgad +5 authors

LLMs encode more internal factual knowledge than they express externally, with some knowledge so deeply hidden that it is never generated, despite repeated sampling.

56large language modelsLLMsHF ↗arXiv ↗
587

Transformer^2: Self-adaptive LLMs

Qi Sun, Edoardo Cetin, Yujin Tang

A self-adaptive framework for large language models uses reinforcement learning to dynamically adjust task-specific components during inference, enhancing adaptability and performance with efficiency.

55self-adaptive large language models (LLMs)fine-tuningHF ↗arXiv ↗
593

MinMo: A Multimodal Large Language Model for Seamless Voice Interaction

Qian Chen, Yafeng Chen, Yanni Chen +33 authors

MinMo, a multimodal large language model, integrates speech and text processing to achieve state-of-the-art performance in voice comprehension and generation, while enabling full-duplex conversation and instruction-following capabilities.

54multimodal large language modelnative modelsHF ↗arXiv ↗
594

Continuous Diffusion Model for Language Modeling

Jaehyeong Jo, Sung Ju Hwang

A continuous diffusion model for language modeling that leverages the geometry of discrete distributions outperforms existing discrete models and matches autoregressive models in performance.

53diffusion modelsautoregressive modelsHF ↗arXiv ↗
595

Region-Adaptive Sampling for Diffusion Transformers

Ziming Liu, Yifan Yang, Chengruidong Zhang +4 authors

RAS, a novel sampling strategy for diffusion transformers, dynamically adjusts sampling ratios based on regions of focus, achieving speedups in diffusion models with minimal quality loss.

53diffusion modelssampling strategyHF ↗arXiv ↗
598

Towards an AI co-scientist

Juraj Gottweis, Wei-Hung Weng, Alexander Daryin +31 authors

A multi-agent AI system named AI co-scientist aids in scientific discovery by generating and validating novel hypotheses across biomedical areas, demonstrating potential improvements in drug repurposing, target discovery, and bacterial evolution understanding.

53multi-agent systemGemini 2.0HF ↗arXiv ↗
20 / 21

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号