CommentsI found a big mistake in the paper that causes significant bias on the results. The residual links are not taken into consideration when computing the transmission. All results about the compressed data size and transmission latency would be affected
机构
*
Yale University(耶鲁大学)
;
Broad Institute of MIT and Harvard(麻省理工学院-哈佛大学博德研究所)
;
Google DeepMind(谷歌DeepMind)
;
Stanford University(斯坦福大学)
;
Genentech(基因泰克)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
Cornell University(康奈尔大学)
;
Harvard University(哈佛大学)
HybridKV: Hybrid KV Cache Compression for Efficient Multimodal Large Language Model Inference
HybridKV: 为高效多模态大语言模型推理设计的混合KV缓存压缩
Bowen Zeng, Feiyang Ren, Jun Zhang, Xiaoling Gu, Ke Chen, Lidan Shou, Huan Li
机构
*
The State Key Laboratory of Blockchain and Data Security, Zhejiang University(浙江大学区块链与数据安全国家重点实验室)
;
Hangzhou High-Tech Zone (Binjiang) Institute of Blockchain and Data Security(杭州高新区(滨江)区块链与数据安全研究院)
;
Hangzhou Dianzi University(杭州电子科技大学)
专题命中
视觉推理
:multimodal large language model(title,abstract);MLLM(abstract);分类 cs.AI
PanopticQuery: Unified Query-Time Reasoning for 4D Scenes
PanopticQuery:4D场景的统一查询时推理
Ruilin Tang, Yang Zhou, Zhong Ye, Wenxi Liu, Yan Huang, Shengfeng He
机构
*
School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院)
;
School of Computing and Information Systems, Singapore Management University(新加坡管理大学计算与信息系统学院)
;
School of Computer Science and Technology, Guangdong University of Technology(广东工业大学计算机科学与技术学院)
;
College of Computer and Data Science, Fuzhou University(福州大学计算机与数据科学学院)
CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment
CoEnv:通过组合环境驱动具身体验多智能体协作
Li Kang, Yutao Fan, Rui Li, Heng Zhou, Yiran Qin, Zhemeng Zhang, Songtao Huang, Xiufeng Song, Zaibin Zhang, Bruno N. Y. Chen, Zhenfei Yin, Dongzhan Zhou, Wangmeng Zuo, Lei Bai
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
University of Science and Technology of China(中国科学技术大学)
;
CUHK-Shenzhen(香港中文大学(深圳))
;
Fudan University(复旦大学)
;
Dalian University of Technology(大连理工大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of Oxford(牛津大学)
TeamPath: Building MultiModal Pathology Experts with Reasoning AI Copilots
TeamPath: 构建多模态病理专家的推理AI助手
Tianyu Liu, Weihao Xuan, Hao Wu, Peter Humphrey, Marcello DiStasio, Mohamed Kahila, Alfonso Garcia Tan, Heli Qi, Rui Yang, Simeng Han, Tinglin Huang, Fang Wu, Chen Liu, Qingyu Chen, Nan Liu, Irene Li, Hua Xu, Hongyu Zhao
机构
*
Interdepartmental Program in Computational Biology and Biomedical Informatics, Yale University(耶鲁大学计算生物学与生物医学信息学跨学科项目)
;
Department of Biostatistics, Yale University(耶鲁大学生物统计学系)
;
Broad Institute of MIT and Harvard(博德研究所)
;
Department of Complexity Science and Engineering, The University of Tokyo(东京大学复杂科学与工程系)
;
Center for Advanced Intelligence Project, RIKEN(理化学研究所先进智能项目中心)
;
Department of Pathology, Yale University(耶鲁大学病理学系)
;
Department of Anatomical Pathology, Singapore General Hospital(新加坡中央医院解剖病理学系)
;
Center for Biomedical Data Science, Duke–NUS Medical School, Singapore, Singapore(杜克-新加坡国立大学医学院生物医学数据科学中心)
;
Department of Computer Science, Yale University(耶鲁大学计算机科学系)
;
Department of Computer Science, Stanford University(斯坦福大学计算机科学系)
专题命中
视觉推理
:visual language model(abstract);分类 cs.CV
机构
*
Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身智能研究所)
;
Shanghai Innovation Institute(上海创新研究院)
;
Shanghai Key Laboratory of Multimodal Embodied AI(上海市多模态具身人工智能重点实验室)
;
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
The Chinese University of Hong Kong(香港中文大学)
;
Central South University(中南大学)
;
Fudan University(复旦大学)
机构
*
School of Mathematics, Shandong University(山东大学数学学院)
;
Shenzhen Loop Area Institute(深圳河套学院)
;
Yeshiva University(叶史瓦大学)
;
Qilu University of Technology(齐鲁工业大学)
;
Shandong Normal University(山东师范大学)
;
Chuzhou University(滁州学院)
Watch Before You Answer: Learning from Visually Grounded Post-Training
在回答前观看:从视觉引导的后训练中学习
Yuxuan Zhang, EunJeong Hwang, Huaisong Zhang, Penghui Du, Yiming Jia, Dongfu Jiang, Xuan He, Shenhui Zhang, Ping Nie, Peter West, Kelsey R. Allen
机构
*
University of British Columbia(不列颠哥伦比亚大学)
;
Vector Institute(向量研究所)
;
Etude AI
;
Kolors Team, Kuaishou Technology(快手科技Kolors团队)
;
University of Toronto(多伦多大学)
;
University of Waterloo(滑铁卢大学)
;
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)