TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Jul 14 – Jul 20, 2025
本周最热263

A Survey of Context Engineering for Large Language Models

Lingrui Mei, Jiayu Yao, Yuyao Ge +12 authors

Context Engineering systematically optimizes information payloads for Large Language Models, addressing gaps in generating sophisticated, long-form outputs.

Large Language ModelsContext Engineeringcontext retrievalcontext generationHF ↗arXiv ↗

50 篇论文 · 按点赞排序

02

Test-Time Scaling with Reflective Generative Model

Zixiao Wang, Yuxin Wang, Xiaorui Wang +8 authors

MetaStone-S1, a reflective generative model using a self-supervised process reward model, achieves high performance with reduced parameters and supports test time scaling.

108reflective generative modelself-supervised process reward modelHF ↗arXiv ↗
09

π^3: Scalable Permutation-Equivariant Visual Geometry Learning

Yifan Wang, Jianjun Zhou, Haoyi Zhu +7 authors

A permutation-equivariant neural network, $\pi^3$, reconstructs visual geometry without a fixed reference view, achieving state-of-the-art performance in camera pose estimation, depth estimation, and point map reconstruction.

67feed-forward neural networkpermutation-equivariant architectureHF ↗arXiv ↗
17

PhysX: Physical-Grounded 3D Asset Generation

Ziang Cao, Zhaoxi Chen, Linag Pan +1 authors

PhysX addresses the gap in physical-grounded 3D asset generation by introducing PhysXNet, a physics-annotated dataset, and PhysXGen, a feed-forward framework that integrates physical knowledge into 3D generation.

45physics-grounded 3D asset generationPhysXNetHF ↗arXiv ↗
21

Scaling Laws for Optimal Data Mixtures

Mustafa Shukor, Louis Bethune, Dan Busbridge +4 authors

Scaling laws predict optimal data mixtures for large foundation models across different domains, improving performance and reducing trial-and-error.

39scaling lawslarge language modelHF ↗arXiv ↗
25

Voxtral

Alexander H. Liu, Andy Ehrenberg, Andy Lo +103 authors

Voxtral Mini and Voxtral Small are multimodal audio chat models that excel in understanding spoken audio and text, with a 32K context window for extended audio and conversation handling.

35multimodal audio chat modelsspoken audioHF ↗arXiv ↗
27

One Token to Fool LLM-as-a-Judge

Yulai Zhao, Haolin Liu, Dian Yu +3 authors

Generative reward models using large language models are vulnerable to superficial manipulations, leading to false positive rewards, and a new data augmentation strategy improves their robustness.

32generative reward modelsLLMs-as-judgesHF ↗arXiv ↗
29

Seq vs Seq: An Open Suite of Paired Encoders and Decoders

Orion Weller, Kathryn Ricci, Marc Marone +3 authors

The Ettin suite of models demonstrates that encoder-only and decoder-only architectures perform optimally in their respective tasks, with encoder-only models excelling at classification and retrieval, and decoder-only models at generation, and that adapting models to different tasks through continued training is less effective.

28large language modeldecoder-only language modelsHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号