TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Mar 25 – Mar 31, 2024
本周最热82

The Unreasonable Ineffectiveness of the Deeper Layers

Andrey Gromov, Kushal Tirumala, Hassan Shapourian +2 authors

Layer pruning of pre-trained LLMs with parameter-efficient finetuning methods shows minimal performance degradation and significant resource savings in both finetuning and inference.

layer-pruningopen-weight pretrained LLMsparameter-efficient finetuningPEFTHF ↗arXiv ↗

45 篇论文 · 按点赞排序

02

LLM Agent Operating System

Kai Mei, Zelong Li, Shuyuan Xu +3 authors

AIOS, an operating system embedding large language models, addresses resource allocation, context switching, and concurrency challenges for intelligent agents, demonstrating reliability and efficiency.

73large language modelintelligent agentsHF ↗arXiv ↗
03

ViTAR: Vision Transformer with Any Resolution

Qihang Fan, Quanzeng You, Xiaotian Han +5 authors

ViTAR enhances Vision Transformers' scalability across resolutions through dynamic token integration and fuzzy positional encoding, improving accuracy and reducing computational costs.

56Vision Transformersdynamic resolution adjustmentHF ↗arXiv ↗
05

sDPO: Don't Use Your Data All at Once

Dahyun Kim, Yungi Kim, Wonho Song +4 authors

A stepwise direct preference optimization approach improves the alignment of large language models with human preferences and enhances their performance.

41large language modelsLLMHF ↗arXiv ↗
06

InternLM2 Technical Report

Zheng Cai, Maosong Cao, Haojiong Chen +97 authors

InternLM2 is an open-source LLM that outperforms predecessors through innovative pre-training and optimization techniques, including Supervised Fine-Tuning and Conditional Online Reinforcement Learning from Human Feedback.

34Large Language ModelsLLMsHF ↗arXiv ↗
08

Can large language models explore in-context?

Akshay Krishnamurthy, Keegan Harris, Dylan J. Foster +2 authors

Experimentation with LLMs in multi-armed bandit environments reveals that robust exploration requires specific prompts, external summarization of history, or algorithmic interventions.

31Large Language ModelsLLMsHF ↗arXiv ↗
11

Long-form factuality in large language models

Jerry Wei, Chengrun Yang, Xinying Song +8 authors

SAFE, a novel method using LLMs and Google Search, evaluates long-form factual accuracy in LLM responses more cheaply and accurately than human annotators.

26LLMsLarge language modelsHF ↗arXiv ↗
13

Garment3DGen: 3D Garment Stylization and Texture Generation

Nikolaos Sarafianos, Tuur Stuyck, Xiaoyu Xiang +3 authors

Garment3DGen synthesizes 3D textile assets from 2D images using diffusion methods and mesh deformation, enabling simulation-ready garments from real or generated images, including textual prompts.

24image to 3D diffusion methodsmesh deformation optimizationHF ↗arXiv ↗
18

LITA: Language Instructed Temporal-Localization Assistant

De-An Huang, Shijia Liao, Subhashree Radhakrishnan +4 authors

LITA addresses temporal localization in multimodal LLMs by introducing time tokens, SlowFast architecture, and a new task dataset, improving video-based text generation and temporal mIoU.

19multimodal LLMstime representationHF ↗arXiv ↗
23

TC4D: Trajectory-Conditioned Text-to-4D Generation

Sherwin Bahmani, Xian Liu, Yifan Wang +9 authors

TC4D, a trajectory-conditioned text-to-4D generation method, improves motion realism and synthesis by separating global motion into rigid transformations and learning local deformations using text-to-video supervision.

17trajectory-conditioned text-to-4D generationrigid transformationHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号