TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Feb 24 – Mar 2, 2025

50 篇论文 · 按点赞排序

31

LightThinker: Thinking Step-by-Step Compression

Jintian Zhang, Yuqi Zhu, Mengshu Sun +6 authors

LightThinker improves efficiency in LLMs by dynamically compressing intermediate thoughts, reducing memory usage and inference time while maintaining performance.

31Large language modelsLightThinkerHF ↗arXiv ↗
34

Rank1: Test-Time Compute for Reranking in Information Retrieval

Orion Weller, Kathryn Ricci, Eugene Yang +3 authors

Rank1 is a re-ranking model using test-time compute and reasoning language models to enhance performance and explainability compared to smaller models, demonstrating state-of-the-art results on reasoning and instruction-following datasets.

29reranking modeltest-time computeHF ↗arXiv ↗
46

MoBA: Mixture of Block Attention for Long-Context LLMs

Enzhe Lu, Zhejun Jiang, Jingyuan Liu +22 authors

Mixture of Block Attention (MoBA) improves large language models' efficiency and performance on long-context tasks by dynamically switching between full and sparse attention without predefined biases.

20Mixture of Block AttentionMoBAHF ↗arXiv ↗
2 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号