TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Nov 18 – Nov 24, 2024

50 篇论文 · 按点赞排序

31

Xmodel-1.5: An 1B-scale Multilingual LLM

Wang Qun, Liu Yang, Lin Qingquan +1 authors

Xmodel-1.5, a 1-billion-parameter multilingual model pretrained on 2 trillion tokens, achieves strong performance across multiple languages and contributes a Thai evaluation dataset.

141-billion-parametermultilingual large modelHF ↗arXiv ↗
32

Number it: Temporal Grounding Videos like Flipping Manga

Yongliang Wu, Xinting Hu, Yuyang Sun +5 authors

Number-Prompt (NumPro) enhances Video Large Language Models (Vid-LLMs) for Video Temporal Grounding (VTG) by adding numerical identifiers to video frames, improving performance significantly without increased computational costs.

14Vid-LLMsVideo Temporal GroundingHF ↗arXiv ↗
38

SlimLM: An Efficient Small Language Model for On-Device Document Assistance

Thang M. Pham, Phat T. Nguyen, Seunghyun Yoon +3 authors

SlimLM, a series of small language models optimized for document assistance tasks on mobile devices, demonstrates efficient performance and enhanced capabilities within size and time constraints, providing a benchmark for on-device language models.

12small language modelsdocument assistance tasksHF ↗arXiv ↗
40

Building Trust: Foundations of Security, Safety and Transparency in AI

Huzaifa Sidhpurwala, Garth Mollett, Emily Fox +2 authors

This paper explores the rapidly evolving ecosystem of publicly available AI models, and their potential implications on the security and safety landscape. As AI models become increasingly prevalent, understanding their potential risks and vulnerabilities is crucial. We review the current security and safety scenarios while highlighting challenges such as tracking issues, remediation, and the apparent absence of AI model lifecycle and ownership processes. Comprehensive strategies to enhance security and safety for both model developers and end-users are proposed. This paper aims to provide some of the foundational pieces for more standardized security, safety, and transparency in the development and operation of AI models and the larger open ecosystems and communities forming around them.

10HF ↗arXiv ↗
41

Adaptive Decoding via Latent Preference Optimization

Shehzaad Dhuliawala, Ilia Kulikov, Ping Yu +4 authors

Adaptive Decoding with Latent Preference Optimization dynamically adjusts sampling temperature during model inference, improving performance across tasks requiring varied temperature settings.

10Adaptive DecodingLatent Preference OptimizationHF ↗arXiv ↗
46

Soft Robotic Dynamic In-Hand Pen Spinning

Yunchao Yao, Uksang Yoo, Jean Oh +2 authors

SWIFT, a system for learning dynamic pen spinning tasks with a soft robotic hand, achieves high success rates using real-world data without prior object knowledge, highlighting the potential for soft robotics in dynamic manipulation.

9soft robotic systemsdynamic in-hand manipulationHF ↗arXiv ↗
2 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号