TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

May 8 – May 14, 2023
本周最热34

StarCoder: may the source be with you!

Raymond Li, Loubna Ben Allal, Yangtian Zi +64 authors

StarCoder, a 15.5B parameter LLM trained on 1 trillion tokens, outperforms other open Code LLMs across multiple languages and fine-tuned Python, with safety enhancements and publicly available under the Open Responsible AI Model license.

Large Language ModelsCode LLMsStarCoderStarCoderBaseHF ↗arXiv ↗

50 篇论文 · 按点赞排序

03

Recommender Systems with Generative Retrieval

Shashank Rajput, Nikhil Mehta, Anima Singh +10 authors

A generative retrieval model using Semantic IDs and a Transformer sequence-to-sequence approach improves recommendation system performance and generalization, especially for cold-start items.

9dual-encoder modelApproximate Nearest NeighborHF ↗arXiv ↗
10

Exploiting Diffusion Prior for Real-World Image Super-Resolution

Jianyi Wang, Zongsheng Yue, Shangchen Zhou +2 authors

A novel method leverages pre-trained text-to-image diffusion models for blind super-resolution, using a time-aware encoder and a controllable feature wrapping module to enhance fidelity and adaptability to varying resolutions.

6text-to-image diffusion modelsblind super-resolutionHF ↗arXiv ↗
18

An Inverse Scaling Law for CLIP Training

Xianhang Li, Zeyu Wang, Cihang Xie

Reducing the token length in CLIP training improves scaling efficiency, allowing academic researchers to train high-performing models with limited resources.

3CLIPfoundation modelHF ↗arXiv ↗
19

VideoChat: Chat-Centric Video Understanding

KunChang Li, Yinan He, Yi Wang +6 authors

VideoChat combines video foundation models and large language models with a learnable neural interface for tasks such as spatiotemporal reasoning and causal relationship inference, supported by a new video-centric instruction dataset.

3video foundation modelslarge language modelsHF ↗arXiv ↗
22

Controllable Light Diffusion for Portraits

David Futschik, Kelvin Ritland, James Vecore +5 authors

A novel learning-based method improves lighting in portraits by softening shadows and specular highlights while preserving overall illumination, and enhances higher-level vision applications.

3light diffusionhashingHF ↗arXiv ↗
28

NerfAcc: Efficient Sampling Accelerates NeRFs

Ruilong Li, Hang Gao, Matthew Tancik +1 authors

NerfAcc accelerates Neural Radiance Field training and rendering by providing advanced, flexible sampling methods that reduce training time significantly.

2Neural Radiance Fieldsvolume renderingHF ↗arXiv ↗
30

Code Execution with Pre-trained Language Models

Chenxiao Liu, Shuai Lu, Weizhu Chen +5 authors

CodeExecutor, a Transformer model enhanced with code execution pre-training and curriculum learning, demonstrates improved performance in code execution tasks and benefits for code intelligence tasks like zero-shot search and text-to-code generation.

2Transformer modelcode execution pre-trainingHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号