ChatPaper.ai
打開菜單
首頁
每日論文
arXiv
HuggingFace
定價
賬戶
工作台
🇭🇰
繁體中文
Loading...
•
•
•
•
•
•
•
•
•
•
AI研究論文每日精選
每日精選AI研究論文及翻譯
February 29th, 2024
1比特LLM時代:所有大型語言模型都在1.58比特。
The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits
Shuming Ma, Hongyu Wang, Lingxiao Ma, Lei Wang, Wenhui Wang, Shaohan Huang, Li Dong, Ruiping Wang, Jilong Xue, Furu Wei
•
Feb 27, 2024
•
618
143
EMO:情感肖像活現 - 在弱條件下利用音訊到影片擴散模型生成具表現力的肖像影片
EMO: Emote Portrait Alive - Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions
Linrui Tian, Qi Wang, Bang Zhang, Liefeng Bo
•
Feb 27, 2024
•
196
20
Sora:關於大視覺模型的背景、技術、限制和機遇的綜述
Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models
Yixin Liu, Kai Zhang, Yuan Li, Zhiling Yan, Chujie Gao, Ruoxi Chen, Zhengqing Yuan, Yue Huang, Hanchi Sun, Jianfeng Gao, Lifang He, Lichao Sun
•
Feb 27, 2024
•
89
5
OmniACT:用於啟用桌面和網頁多模式通用自主代理的數據集和基準。
OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web
Raghav Kapoor, Yash Parag Butala, Melisa Russak, Jing Yu Koh, Kiran Kamble, Waseem Alshikh, Ruslan Salakhutdinov
•
Feb 27, 2024
•
26
6
當擴展遇上LLM微調:資料、模型和微調方法的影響
When Scaling Meets LLM Finetuning: The Effect of Data, Model and Finetuning Method
Biao Zhang, Zhongtao Liu, Colin Cherry, Orhan Firat
•
Feb 27, 2024
•
26
3
無需訓練的大型語言模型長文本擴展
Training-Free Long-Context Scaling of Large Language Models
Chenxin An, Fei Huang, Jun Zhang, Shansan Gong, Xipeng Qiu, Chang Zhou, Lingpeng Kong
•
Feb 27, 2024
•
25
4
DiffuseKronA:一種用於個性化擴散模型的參數高效微調方法
DiffuseKronA: A Parameter Efficient Fine-tuning Method for Personalized Diffusion Model
Shyam Marjit, Harshit Singh, Nityanand Mathur, Sayak Paul, Chia-Mu Yu, Pin-Yu Chen
•
Feb 27, 2024
•
25
1
影片作為現實世界決策的新語言
Video as the New Language for Real-World Decision Making
Sherry Yang, Jacob Walker, Jack Parker-Holder, Yilun Du, Jake Bruce, Andre Barreto, Pieter Abbeel, Dale Schuurmans
•
Feb 27, 2024
•
22
1
評估LLM智能體的非常長期對話記憶
Evaluating Very Long-Term Conversational Memory of LLM Agents
Adyasha Maharana, Dong-Ho Lee, Sergey Tulyakov, Mohit Bansal, Francesco Barbieri, Yuwei Fang
•
Feb 27, 2024
•
20
3
邁向語言模型的最佳學習
Towards Optimal Learning of Language Models
Yuxian Gu, Li Dong, Yaru Hao, Qingxiu Dong, Minlie Huang, Furu Wei
•
Feb 27, 2024
•
18
1
Sora以令人驚嘆的幾何一致性生成視頻。
Sora Generates Videos with Stunning Geometrical Consistency
Xuanyi Li, Daquan Zhou, Chenxu Zhang, Shaodong Wei, Qibin Hou, Ming-Ming Cheng
•
Feb 27, 2024
•
18
1
視覺與聽覺:使用擴散潛在對齊器進行開放域視聽生成
Seeing and Hearing: Open-domain Visual-Audio Generation with Diffusion Latent Aligners
Yazhou Xing, Yingqing He, Zeyue Tian, Xintao Wang, Qifeng Chen
•
Feb 27, 2024
•
16
1
Playground v2.5:三個關於提升文本到圖像生成中美學品質的見解
Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation
Daiqing Li, Aleks Kamko, Ehsan Akhgari, Ali Sabet, Linmiao Xu, Suhail Doshi
•
Feb 27, 2024
•
12
1
具有佈局學習的解耦式3D場景生成
Disentangled 3D Scene Generation with Layout Learning
Dave Epstein, Ben Poole, Ben Mildenhall, Alexei A. Efros, Aleksander Holynski
•
Feb 26, 2024
•
12
1
VastGaussian:用於大型場景重建的大型3D高斯模型
VastGaussian: Vast 3D Gaussians for Large Scene Reconstruction
Jiaqi Lin, Zhihao Li, Xiao Tang, Jianzhuang Liu, Shiyong Liu, Jiayue Liu, Yangdi Lu, Xiaofei Wu, Songcen Xu, Youliang Yan, Wenming Yang
•
Feb 27, 2024
•
11
45