ERASE: Eliminating Redundant Visual Tokens via Adaptive Two-Stage Token Pruning
ERASE: 通过自适应两阶段令牌剪枝消除冗余视觉令牌
Yuna Lee, Kyoungho Min, Yulhwa Kim
机构
*
Department of Electrical and Computer Engineering, Sungkyunkwan University, Republic of Korea(电气与计算机工程系,成均馆大学,大韩民国)
;
Department of Semiconductor Systems Engineering, Sungkyunkwan University, Republic of Korea(半导体系统工程系,成均馆大学,大韩民国)
专题命中
效率与部署
:large language model(abstract);language model(abstract)
机构
*
School of Mathematics and Statistics(数学与统计学学院)
;
Xi’an Jiaotong University(西安交通大学)
;
Gaoling School of Artificial Intelligence(白洋学校人工智能学院)
;
Renmin University of China(中国人民大学)
机构
*
School of Computing and Data Science(计算与数据科学学院)
;
The University of Hong Kong(香港大学)
;
School of Electrical Engineering and Computer Science(电气工程与计算机科学学院)
;
The University of Queensland(昆士兰大学)
;
Department of Computing(计算学院)
机构
*
EPIC Lab, Shanghai Jiao Tong University(上海交通大学EPIC实验室)
;
Tsinghua University(清华大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Fudan University(复旦大学)
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models
AR-VLA:面向视觉-语言-动作模型的真自回归动作专家
Yutong Hu, Jan-Nico Zaech, Nikolay Nikolov, Yuanqi Yao, Sombit Dey, Giuliano Albanese, Renaud Detry, Luc Van Gool, Danda Paudel
机构
*
KU Leuven, Dept. Mechanical Engineering, Research unit Robotics, Automation and Mechatronics(库勒恩大学,机械工程系,机器人、自动化与机电一体化研究单位)
;
KU Leuven, Dept. Electrical Engineering, Research unit Processing Speech and Images(库勒恩大学,电气工程系,语音和图像处理研究单位)
Optimal Attention Temperature Improves the Robustness of In-Context Learning under Distribution Shift in High Dimensions
最优注意力温度提升高维分布偏移下的上下文学习鲁棒性
Samet Demir, Zafer Dogan
机构
*
MLIP Research Group, KUIS AI Center, Koç University(MLIP研究组、KUIS人工智能中心、科克大学)
;
Department of EEE, Koç University, İstanbul, Turkey(电子工程系、科克大学、伊斯坦布尔,土耳其)
Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring
Arcane:通过语义聚类和MCTS引导的规则探索的断言减少框架
Hongqin Lyu, Yonghao Wang, Zhiteng Chao, Tiancheng Wang, Huawei Li
机构
*
State Key Lab of Processors, Institute of Computing Technology, CAS, Beijing, China(处理器国家重点实验室,计算技术研究所,中国科学院,北京,中国)
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)
Language-Conditioned Visual Grounding with CLIP Multilingual
基于CLIP多语言的语言条件视觉 grounding
J. de Curtò, Mauro Liz, I. de Zarzà
机构
*
1 Department of Computer Applications in Science \& Engineering, BARCELONA Supercomputing Center, Barcelona, Spain
;
3 Department of Electrical
;
Computer Engineering, Boston University, Boston, MA 02215, USA
;
4 Human centered AI, Data \& Software, LUXEMBOURG Institute of Science
Anchoring the Eigengap: Cross-Modal Spectral Stabilization for Sample-Efficient Representation Learning
锚定特征间隙:跨模态谱稳定化以实现样本高效表征学习
Nikhil J. Dhinagar, Vidhi Chhatbar, Chirag Jagad, Pavithra Senthilkumar, Sophia I. Thomopoulos, Mahir H. Khan, Sook-Lei Liew, the ENIGMA-Stroke Recovery Working Group, Paul M. Thompson
机构
*
Imaging Genetics Center, Mark & Mary Stevens Neuroimaging & Informatics Institute, Keck School of Medicine, University of Southern California(影像基因中心,马克与玛丽史蒂文斯神经影像与信息学研究所,凯克医学院,南加州大学)
;
Neuroscience Graduate Program, Mark & Mary Stevens Neuroimaging & Informatics Institute, Chan Division of Occupational Science & Occupational Therapy, Biomedical Engineering, University of Southern California(神经科学研究生项目,马克与玛丽史蒂文斯神经影像与信息学研究所,查恩职业科学与职业治疗 division,生物医学工程,南加州大学)
机构
*
Zhejiang University(浙江大学)
;
Nanyang Technological University(南洋理工大学)
;
Sun Yat-sen University(中山大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Virginia Tech(弗吉尼亚理工大学)