Grad-ECLIP: Gradient-based Visual and Textual Explanations for CLIP
Grad-ECLIP: 基于梯度的CLIP视觉与文本解释
Chenyang Zhao, Kun Wang, Janet H. Hsiao, Antoni B. Chan
机构
*
Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系)
;
Division of Social Science and Department of Computer Science & Engineering, Hong Kong University of Science & Technology(香港科学与技术大学社会科学学院及计算机科学与工程系)
;
SenseTime Group Ltd(时光集团有限公司)
Journal refZhao C, Wang K, Hsiao J H, et al. Grad-eclip: Gradient-based visual and textual explanations for clip[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2026
Comments31 pages, 15 tables. Accepted by IEEE Transactions on Pattern Analysis and Machine Intelligence. The supplementary material is included at the end of the manuscript
DeepForgeSeal: Latent Space-Driven Semi-Fragile Watermarking for Deepfake Detection Using Adversarial Reinforcement Learning
DeepForgeSeal:基于对抗强化学习的隐空间驱动半脆弱深度伪造检测水印技术
Tharindu Fernando, Clinton Fookes, Sridha Sridharan
机构
*
The Signal Processing, Artificial Intelligence and Vision Technologies (SAIVT), Queensland University of Technology, Australia(信号处理、人工智能与视觉技术研究所(SAIVT),昆士兰理工大学)
BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation
BridgeVLA++:一种面向三维操作的数据高效、可泛化且内存增强的视觉-语言-动作框架
Peiyan Li, Yuze Zhu, Yixiang Chen, Qisen Ma, Yuan Xu, Jiabing Yang, He Guan, Yan Huang, Hongtao Wu, Xiao Ma, Tao Kong, Liang Wang, Tieniu Tan
机构
*
New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室(NLPR))
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
FiveAges
;
ByteDance Seed(字节跳动种子实验室)
CommentsThis work has been submitted to the IEEE TPAMI for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible
Towards Ultrafast Depth Sensing Via Active Event-based Stereo Vision
通过基于主动事件的立体视觉实现超快速深度感知
Jianing Li, Yunjian Zhang, Haiqian Han, Kangyao Huang, Xiangyang Ji
机构
*
School of Computer Science, Peking University(北京大学计算机科学学院)
;
Peng Cheng Laboratory(鹏城实验室)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
机构
*
National Key Laboratory of Multispectral Information Intelligent Processing Technology, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院多光谱信息智能处理技术国家重点实验室)
;
School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)
机构
*
Institute of Artificial Intelligence, State Key Laboratory of Virtual Reality Technology and Systems, Beihang University(北京航空航天大学虚拟现实技术与系统国家重点实验室人工智能研究院)
;
College of Artificial Intelligence, Tsinghua University(清华大学人工智能学院)
;
Security Department, Alibaba Group(阿里巴巴集团安全部)
;
School of Cyber Science and Technology, Shenzhen Campus of Sun Yat-Sen University(中山大学深圳校区网络科学与技术学院)
机构
*
Guangdong Provincial Key Laboratory of Visual Media and Multidimensional Intelligence, CSSE, Shenzhen University, China(广东省视觉媒体与多维智能重点实验室,计算机科学与电子技术学院,深圳大学,中国)
;
Fujian Key Laboratory of Sensing and Computing for Smart City, School of Informatics, Xiamen University, China(福建省智慧城市感知与计算重点实验室,信息学院,厦门大学,中国)
;
Department of Computer Science, Hong Kong Baptist University, China(计算机科学系,香港 Baptist 大学,中国)
;
Inspur Smart City Technology Co., Ltd., China(Inspur 智能城市技术有限公司,中国)
机构
*
School of Artificial Intelligence (SAI), Shanghai Jiao Tong University(人工智能学院(SAI),上海交通大学)
;
School of Mechanical Engineering, Shanghai Jiao Tong University(机械工程学院,上海交通大学)
;
School of Electronic Information and Electrical Engineering, Shanghai Jiao Tong University(电子信息与电气工程学院,上海交通大学)
;
Institute of Automation Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所(CASIA))
Jiahong Zhang, Sijun Shen, Man Yao, Han Xu, Mingqiang Huang, Yonghong Tian, Bo Xu, Guoqi Li
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
State Key Laboratory of Media Convergence and Communication, Communication University of China(中国传媒大学媒体融合与传播国家重点实验室)
;
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
Peng Cheng Laboratory(鹏城实验室)
;
Institute for Artificial Intelligence, Peking University(北京大学人工智能研究院)
机构
*
School of Computer Science and Engineering, Southeast University(计算机科学与工程学院,东南大学)
;
Key Laboratory of Computer Network and Information Integration (Southeast University), MoE, China(计算机网络与信息集成重点实验室(东南大学),教育部,中国)
;
School of Information and Physical Sciences, The University of Newcastle(信息与物理科学学院,新castle大学)
SWIFT: A Small-World Interaction Framework for Flow-Aware Trajectory Prediction in Autonomous Driving
SWIFT:用于自动驾驶中流量感知轨迹预测的小世界交互框架
Chengyue Wang, Bin Rao, Haicheng Liao, Bonan Wang, Chengzhong Xu, Zhenning Li
机构
*
University of Macau(澳门大学)
;
State Key Laboratory of Internet of Things for Smart City, University of Macau(澳门大学智慧城市物联网国家重点实验室)
;
Department of Civil and Environmental Engineering, University of Macau(澳门大学土木与环境工程系)
机构
*
School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院)
;
Institute of Big Data, Fudan University(复旦大学大数据研究院)
;
AI Thrust, Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)人工智能方向)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算机与数据科学学院)
Benchmarking the Robustness of Autonomous Driving to Environmental Illusions: A Lane Perception Perspective
从车道感知角度评估自动驾驶对环境错觉的鲁棒性:基准测试
Tianyuan Zhang, Xianglong Liu, Aishan Liu, Lu Wang, Yitong Zhang, Peng Yue, Mingchuan Zhang, Siyuan Liang, Dacheng Tao
机构
*
SKLCCSE, the School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院软件安全技术与工程北京市重点实验室)
;
the School of Cyber Science and Technology, Sun Yat-sen University(中山大学网络空间科学与技术学院)
;
Henan University of Science and Technology(河南科技大学)
;
the School of Computing, National University of Singapore(新加坡国立大学计算学院)
;
College of Computing & Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)