TensorX

Explore · 每周精选

发现最受关注的研究论文,追踪研究趋势,订阅感兴趣的期刊与关键词。

Nov 4 – Nov 10, 2024
本周最热128

OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models

Siming Huang, Tianhao Cheng, Jason Klein Liu +16 authors

OpenCoder is a top-tier open-access code LLM that provides comprehensive reproducible resources, including data, pipelines, and protocols, to advance code AI research.

large language modelscode generationreasoning tasksagent systemsHF ↗arXiv ↗

50 篇论文 · 按点赞排序

04

BitNet a4.8: 4-bit Activations for 1-bit LLMs

Hongyu Wang, Shuming Ma, Furu Wei

BitNet a4.8 enhances the efficiency of large language models through 4-bit quantization and sparsification, achieving equivalent performance to BitNet b1.58 with reduced inference costs.

701-bit Large Language Models (LLMs)BitNet b1.58HF ↗arXiv ↗
14

How Far is Video Generation from World Model: A Physical Law Perspective

Bingyi Kang, Yang Yue, Rui Lu +5 authors

Diffusion-based video generation models demonstrate perfect in-distribution generalization and measurable scaling behavior for combinatorial generalization but fail in out-of-distribution scenarios, prioritizing color over other factors.

34diffusion-based video generation modelsin-distributionHF ↗arXiv ↗
15

Personalization of Large Language Models: A Survey

Zhehao Zhang, Ryan A. Rossi, Branislav Kveton +18 authors

A taxonomy is introduced to bridge the gap between personalized text generation and personalization in LLMs for downstream applications, providing a comprehensive view of personalization techniques, datasets, and challenges.

33Large Language Models (LLMs)personalized text generationHF ↗arXiv ↗
21

Analyzing The Language of Visual Tokens

David M. Chan, Rodolfo Corona, Joonyong Park +3 authors

Analysis of transformer-based models like LLaVA and Chameleon reveals similarities and differences in statistical behavior between discrete visual languages and natural languages, highlighting directions for improving computer vision models.

24transformer-based modelsLLaVAHF ↗arXiv ↗
23

MVPaint: Synchronized Multi-View Diffusion for Painting Anything 3D

Wei Cheng, Juncheng Mu, Xianfang Zeng +8 authors

MVPaint, a 3D texturing framework, addresses local discontinuities and multi-view inconsistencies in Text-to-Texture generation by employing synchronized multi-view generation, spatial-aware 3D inpainting, and UV refinement techniques, achieving high-fidelity textures.

24synchronized multi-view generationspatial-aware 3D inpaintingHF ↗arXiv ↗
25

Constant Acceleration Flow

Dogyun Park, Sojin Lee, Sihyeon Kim +3 authors

A novel framework called Constant Acceleration Flow (CAF) improves few-step image generation by incorporating acceleration into the ODE flow model, demonstrating superior performance compared to existing methods.

24Rectified flowreflow proceduresHF ↗arXiv ↗
29

LLaMo: Large Language Model-based Molecular Graph Assistant

Jinyoung Park, Minseong Bae, Dohwan Ko +1 authors

LLaMo, a large language model for molecular graphs, uses a multi-level graph projector and cross-attention to achieve top performance in tasks like molecular description, property prediction, and IUPAC name generation.

22Large Language ModelsLarge Vision-Language ModelsHF ↗arXiv ↗
1 / 2

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号