机构
*
The University of Tokyo(东京大学)
;
Nara Institute of Science and Technology(奈良科学技术研究所)
;
Chungnam National University(忠南国立大学)
;
Institute of Science Tokyo(东京科学大学)
机构
*
Ant Group(蚂蚁集团)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Xi'an Polytechnic University(西安理工大学)
BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation
BridgeVLA++:一种面向三维操作的数据高效、可泛化且内存增强的视觉-语言-动作框架
Peiyan Li, Yuze Zhu, Yixiang Chen, Qisen Ma, Yuan Xu, Jiabing Yang, He Guan, Yan Huang, Hongtao Wu, Xiao Ma, Tao Kong, Liang Wang, Tieniu Tan
机构
*
New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室(NLPR))
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
FiveAges
;
ByteDance Seed(字节跳动种子实验室)
CommentsThis work has been submitted to the IEEE TPAMI for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible
Privacy-Preserving Action Recognition: Taxonomy, Methods, and Privacy-Utility Trade-offs
隐私保护动作识别:分类、方法与隐私-效用权衡
Sareer Ul Amin, Muhammad Ayaz, Muhammad Munsif, Sanghyun Seo
机构
*
Chung-Ang University(中央大学)
;
Ulsan National Institute of Science and Technology (UNIST)(蔚山科学技术院)
;
College of Art and Technology, Chung-Ang University(中央大学艺术与技术学院)
Evaluating the Diagnostic Robustness of Vision-Language Models Under Visual and Textual Perturbations
评估视觉-语言模型在视觉和文本扰动下的诊断鲁棒性
Ali Khoramfar, Mohammad Javad Dousti, Alireza Mohamadian, Heshaam Faili
机构
*
University of Tehran(德黑兰大学)
;
Tehran University of Medical Sciences(德黑兰医科大学)
;
Advanced Diagnostic and Interventional Radiology Research Center (ADIR)(高级诊断与介入放射学研究中心(ADIR))