Device-Cloud Collaborative LLM Inference with Multi-Modal, Multi-Task, Multi-Turn Conversations
具有多模态、多任务、多轮对话的设备-云协作大语言模型推理
Liangqi Yuan, Dong-Jun Han, Shiqiang Wang, Christopher G. Brinton
机构
*
School of Electrical and Computer Engineering, Purdue University(普渡大学电气与计算机工程学院)
;
Department of Computer Science and Engineering, Yonsei University(延世大学计算机科学与工程系)
;
Department of Computer Science, University of Exeter(埃克塞特大学计算机科学系)
专题命中
效率与部署
:LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.LG
EasyOPD: An Easy-to-use On-Policy Distillation Framework for Large Language Models
EasyOPD:一种易于使用的大语言模型在线策略蒸馏框架
Jie Sun, Mao Zheng, Mingyang Song, Qiyong Zhong, Gengsheng Li, Zhepei Hong, Chang Wu, Pengfei Liu, Junfeng Fang, Xiang Wang
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Tencent(腾讯)
;
Shanghai Innovation Institute(上海创新研究院)
;
National University of Singapore(新加坡国立大学)
专题命中
效率与部署
:large language model(title,abstract);language model(title,abstract);分类 cs.CL
机构
*
SAIVT Group, School of Electrical Engineering and Robotics, Queensland University of Technology(澳大利亚昆士兰科技大学电气工程与机器人学院SAIVT组)
;
CSIRO Robotics, CSIRO(澳大利亚联邦科学与工业研究组织机器人部)
机构
*
Faculty of Engineering, The University of Hong Kong(香港大学工程学院)
;
Costello College of Business, George Mason University(乔治·马歇尔大学商学院)
;
Faculty of Business and Economics, The University of Hong Kong(香港大学商学院)
;
College of Engineering, University of California, Berkeley(加州大学伯克利分校工程学院)
专题命中
效率与部署
:foundation model(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
Nested-ReFT:通过离策略展开实现大语言模型微调的高效强化学习
Maxime Heuillet, Yufei Cui, Boxing Chen, Audrey Durand, Prasanna Parthasarathi
机构
*
Mila - Québec AI Institute, Canada(魁北克人工智能研究所)
;
Huawei Noah's Ark Lab (Montreal Research Center), Canada(华为诺亚实验室(蒙特利尔研究中心))
;
Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)
专题命中
效率与部署
:large language model(title);language model(title);post-training(abstract);分类 cs.CL、cs.AI、cs.LG
PromptGraph: Graph-Guided Prompt Sanitization for Balancing Privacy and Utility in LLM Inference
PromptGraph:用于在大语言模型推理中平衡隐私与效用的图引导提示净化
Chen Gu, Hui Wan, Donghui Hu, Hui Wang, Zhuoer Gu
机构
*
School of Computer Science and Information Engineering, Hefei University of Technology(计算机科学与信息工程学院,合肥工业大学)
;
International College Beijing, China Agricultural University(北京国际学院,中国农业大学)
专题命中
效率与部署
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI
机构
*
School of AI, Shanghai Jiao Tong University(上海交通大学人工智能学院)
;
Xiaohongshu Inc.(小红书公司)
;
The Chinese University of Hong Kong(香港中文大学)
;
University of Science and Technology of China(中国科学技术大学)