Visualization of Machine Learning Models through Their Spatial and Temporal Listeners
通过空间和时间监听器可视化机器学习模型
Siyu Wu, Lei Shi, Lei Xia, Cenyang Wu, Zipeng Liu, Yingchaojie Feng, Liang Zhou, Wei Chen
机构
*
School of Computer Science & Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)
;
Institute of Medical Technology, Peking University Health Science Center and National Institute of Health Data Science, Peking University(北京大学医学部医学技术研究院与北京大学健康数据科学国家研究所)
;
School of Software, Beihang University(北京航空航天大学软件学院)
;
National University of Singapore(新加坡国立大学)
;
State Key Laboratory of CAD&CG, Zhejiang University(浙江大学计算机辅助设计与图形学国家重点实验室)
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shanghai Innovation Institute(上海创新研究院)
;
Shanghai Institute of Optics and Fine Mechanics(上海光学精密机械研究所)
;
Fudan University(复旦大学)
;
University of Cambridge(剑桥大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
The University of Hong Kong(香港大学)
;
Fuzhou University(福州大学)
;
University of Washington(华盛顿大学)
;
Stanford University(斯坦福大学)
;
Incept Labs
;
Monash University(莫纳什大学)
;
Ruijin Hospital, Shanghai Jiao Tong University School of Medicine(上海交通大学医学院附属瑞金医院)
;
Alibaba DAMO Academy(阿里巴巴达摩院)
;
The Hong Kong Polytechnic University(香港理工大学)
;
South China University of Technology(华南理工大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Yau Mathematical Sciences Center, Tsinghua University(清华大学丘成桐数学科学中心)
;
Chinese Academy of Sciences(中国科学院)
;
Tsinghua University(清华大学)
;
Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
;
Artificial Intelligence Innovation and Incubation Institute, Fudan University(复旦大学人工智能创新与产业研究院)
;
Shanghai Academy of Artificial Intelligence for Science(上海科学智能研究院)
;
Nankai University(南开大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Zhejiang University(浙江大学)
;
School of Informatics, Xiamen University(厦门大学信息学院)
;
University College London(伦敦大学学院)
;
Sun Yat-sen University(中山大学)
;
Alibaba Group, DAMO Academy, New York, NY, USA(阿里巴巴集团达摩院(纽约))
;
Tongji University(同济大学)
;
University of Toronto(多伦多大学)
;
Department of Psychological and Cognitive Sciences, Tsinghua University(清华大学心理与认知科学系)
;
Johns Hopkins University(约翰霍普金斯大学)
;
Arizona State University(亚利桑那州立大学)
;
Academy for Clinical Innovation and Translation of Shanghai(上海临床创新转化研究院)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
;
ELLIS Institute Finland(芬兰ELLIS研究所)
;
Aalto University(阿尔托大学)
;
Shandong University(山东大学)
;
Xi’an Jiaotong University(西安交通大学)
Dual-Path Learning based on Frequency Structural Decoupling and Regional-Aware Fusion for Low-Light Image Super-Resolution
基于频率结构解耦和区域感知融合的双路径学习用于低光照图像超分辨率
Ji-Xuan He, Jia-Cheng Zhao, Feng-Qi Cui, Jinyang Huang, Yang Liu, Sirui Zhao, Meng Li, Zhi Liu
机构
*
Hefei University of Technology(合肥工业大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Zhejiang University(浙江大学)
;
The University of Electro-Communications(电气通信大学)
Echoes of ownership: Adversarial-guided dual injection for copyright protection in MLLMs
所有权的回声:对抗引导的双注入用于MLLMs中的版权保护
Chengwei Xia, Fan Ma, Ruijie Quan, Yunqiu Xu, Kun Zhan, Yi Yang
机构
*
School of Information Science and Engineering, Lanzhou University(兰州大学信息科学与工程学院)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
DSO: Dual-Scale Neural Operators for Stable Long-term Fluid Dynamics Forecasting
DSO:双尺度神经算子用于稳定长期流体动力学预测
Huanshuo Dong, Hao Wu, Hong Wang, Qin-Yi Zhang, Zhezheng Hao
机构
*
Institute for Clarity in Documentation(文档清晰度研究所)
;
Inria Paris-Rocquencourt(法国国家信息与自动化研究所巴黎-罗康库尔)
;
Rajiv Gandhi University(拉吉夫·甘地大学)
;
Tsinghua University(清华大学)
;
Palmer Research Laboratories(帕尔默研究实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Zhejiang University(浙江大学)
HASS: Hierarchical Simulation of Logopenic Aphasic Speech for Scalable PPA Detection
HASS:面向可扩展PPA检测的层次化日语口语模拟
Harrison Li, Kevin Wang, Cheol Jun Cho, Jiachen Lian, Rabab Rangwala, Chenxu Guo, Emma Yang, Lynn Kurteff, Zoe Ezzes, Willa Keegan-Rodewald, Jet Vonk, Siddarth Ramkrishnan, Giada Antonicelli, Zachary Miller, Marilu Gorno Tempini, Gopala Anumanchipalli
机构
*
UC Berkeley(加州大学伯克利分校)
;
UCSF(加州大学旧金山分校)
;
Zhejiang University(浙江大学)
;
Columbia University(哥伦比亚大学)
;
Basque Center on Cognition, Brain and Language(巴斯克认知、大脑与语言中心)
MA-Bench: Towards Fine-grained Micro-Action Understanding
MA-Bench:迈向细粒度微动作理解
Kun Li, Jihao Gu, Fei Wang, Zhiliang Wu, Hehe Fan, Dan Guo
机构
*
CVLab, College of Information Technology, United Arab Emirates University(阿拉伯联合酋长国大学信息技术学院CVLab)
;
University College London(伦敦大学学院)
;
Hefei University of Technology(合肥工业大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
;
CCAI, Zhejiang University(浙江大学计算机辅助设计与图形学国家重点实验室)
Verify Claimed Text-to-Image Models via Boundary-Aware Prompt Optimization
通过边界感知提示优化验证声称的文本到图像模型
Zidong Zhao, Yihao Huang, Qing Guo, Tianlin Li, Anran Li, Kailong Wang, Jin Song Dong, Geguang Pu
机构
*
Zhejiang University(浙江大学)
;
East China Normal University(华东师范大学)
;
Nankai University(南开大学)
;
Beihang University(北京航空航天大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
National University of Singapore(新加坡国立大学)
机构
*
Southern University of Science and Technology(南方科技大学)
;
Zhejiang University(浙江大学)
;
Technical University of Munich(慕尼黑工业大学)
;
City University of Hong Kong(香港城市大学)
;
Shanghai Jiao Tong University(上海交通大学)
IVEBench: Modern Benchmark Suite for Instruction-Guided Video Editing Assessment
IVEBench:现代指令引导视频编辑评估基准套件
Yinan Chen, Jiangning Zhang, Teng Hu, Yuxiang Zeng, Zhucun Xue, Qingdong He, Chengjie Wang, Yong Liu, Xiaobin Hu, Shuicheng Yan
机构
*
Zhejiang University(浙江大学)
;
Tencent Youtu Lab(腾讯优图实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Auckland(奥克兰大学)
;
National University of Singapore(新加坡国立大学)
Uncovering What, Why and How: A Comprehensive Benchmark for Causation Understanding of Video Anomaly
揭示是什么、为什么和如何:面向视频异常因果理解的全面基准
Hang Du, Sicheng Zhang, Binzhu Xie, Guoshun Nan, Jiayang Zhang, Junrui Xu, Hangyu Liu, Sicong Leng, Jiangming Liu, Hehe Fan, Dajiu Huang, Jing Feng, Linli Chen, Can Zhang, Xuhuan Li, Hao Zhang, Jianhang Chen, Qimei Cui, Xiaofeng Tao
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Queen Mary University of London(伦敦玛丽女王大学)
;
Nanyang Technological University(南洋理工大学)
;
Yunnan University(云南大学)
;
Zhejiang University(浙江大学)
;
China Telecom Co., Ltd. Sichuan Branch(中国电信股份有限公司四川分公司)
机构
*
School of Information Science and Engineering, Hangzhou Normal University(杭州师范大学信息科学与工程学院)
;
State Key Laboratory of Industrial Control and Technology, Zhejiang University(浙江大学工业控制技术国家重点实验室)
;
Neuromorphic Computing and Robotic Cognition Lab, Zhejiang University(浙江大学神经形态计算与机器人认知实验室)
;
Department of Automation, Zhejiang University of Technology(浙江工业大学自动化系)
AD-CARE: A Guideline-grounded, Modality-agnostic LLM Agent for Real-world Alzheimer's Disease Diagnosis with Multi-cohort Assessment, Fairness Analysis, and Reader Study
Wenlong Hou, Sheng Bi, Guangqian Yang, Lihao Liu, Ye Du, Hanxiao Xue, Juncheng Wang, Yuxiang Feng, Yue Xun, Nanxi Yu, Ning Mao, Mo Yang, Yi Wah Eva Cheung, Ling Long, Kay Chen Tan, Lequan Yu, Xiaomeng Ma, Shaozhen Yan, Shujun Wang
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
Xuanwu Hospital, Capital Medical University(首都医科大学宣武医院)
;
Amazon(亚马逊)
;
Zhejiang University(浙江大学)
;
Sany AI(三一人工智能)
;
Yantai Yuhuangding Hospital, Qingdao University(青岛大学烟台毓璜顶医院)
;
The University of Hong Kong(香港大学)
;
The Third Affiliated Hospital of Sun Yat-sen University(中山大学附属第三医院)
Monocular Normal Estimation via Shading Sequence Estimation
单目法线估计 via 阴影序列估计
Zongrui Li, Xinhua Ma, Minghui Hu, Yunqing Zhao, Yingchen Yu, Qian Zheng, Chang Liu, Xudong Jiang, Song Bai
机构
*
School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电气与电子工程学院)
;
ByteDance(字节跳动)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
The State Key Lab of Brain-Machine Intelligence, Zhejiang University(浙江大学脑机智能全国重点实验室)
;
MoE Key Laboratory of Interdisciplinary Research of Computation and Economics, Shanghai University of Finance and Economics(上海财经大学计算与经济学交叉学科教育部重点实验室)
Bohan Jia, Wenxuan Huang, Yuntian Tang, Junbo Qiao, Jincheng Liao, Shaosheng Cao, Fei Zhao, Zhaopeng Feng, Zhouhong Gu, Zhenfei Yin, Lei Bai, Wanli Ouyang, Lin Chen, Fei Zhao, Yao Hu, Zihan Wang, Yuan Xie, Shaohui Lin
机构
*
East China Normal University(华东师范大学)
;
Xiaohongshu Inc.(小红书)
;
KLATASDS, MOE, China(中国教育部统计与数据科学重点实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
Zhejiang University(浙江大学)
;
Fudan University(复旦大学)
;
University of Oxford(牛津大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Nanjing University(南京大学)
CODER: Coupled Diversity-Sensitive Momentum Contrastive Learning for Image-Text Retrieval
CODER: 耦合多样性敏感动量对比学习用于图像-文本检索
Haoran Wang, Dongliang He, Wenhao Wu, Boyang Xia, Min Yang, Fu Li, Yunlong Yu, Zhong Ji, Errui Ding, Jingdong Wang
机构
*
Department of Computer Vision Technology (VIS), Baidu Inc., Beijing, China(百度公司计算机视觉技术部(VIS),北京,中国)
;
Key Lab of Intelligent Information Processing of Chinese Academy of Sciences (CAS), Institute of Computing Technology, CAS, Beijing, China(中国科学院智能信息处理重点实验室,中国科学院计算技术研究所,北京,中国)
;
College of Information Science & Electronic Engineering, Zhejiang University, Hangzhou, China(浙江大学信息与电子工程学院,杭州,中国)
;
School of Electrical & Information Engineering, Tianjin University, Tianjin, China(天津大学电气与信息工程学院,天津,中国)
;
The University of Sydney, Sydney, Australia(悉尼大学,悉尼,澳大利亚)
GIFT: Global Irreplaceability Frame Targeting for Efficient Video Understanding
GIFT:全局不可替代性帧目标用于高效视频理解
Junpeng Ma, Sashuai Zhou, Guanghao Li, Xin Gao, Yue Cao, Hengyu Zeng, Yuxiang Yan, Zhibin Wang, Jun Song, Bo Zheng, Shanghang Zhang, Jian Pu
机构
*
Institute of Science and Technology for Brain-inspired Intelligence, Fudan University(复旦大学类脑智能科学与技术研究院)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(北京大学计算机学院多媒体信息处理国家重点实验室)
;
Zhejiang University(浙江大学)
;
Alibaba Group Holding Limited(阿里巴巴集团控股有限公司)
;
Future Living Lab of Alibaba(阿里巴巴未来生活实验室)
System-Anchored Knee Estimation for Low-Cost Context Window Selection in PDE Forecasting
基于系统锚点的低成本上下文窗口选择方法用于PDE预测
Wenshuo Wang, Fan Zhang
机构
*
School of Future Technology, South China University of Technology(华南理工大学未来技术学院)
;
State Key Laboratory of Ocean Sensing & Ocean College, Zhejiang University(浙江大学海洋学院海洋传感国家重点实验室)
;
Kavli Institute for Astrophysics and Space Research, Massachusetts Institute of Technology(麻省理工学院卡弗里天体物理与空间研究所)
Shopping with a Platform AI Assistant: Who Adopts, When in the Journey, and What For
通过平台AI助手购物:谁采用、何时在旅程中、为何采用
Se Yan, Han Zhong, Zemin, Zhong, Wenyu Zhou
机构
*
Guanghua School of Management, Peking University(北京大学光华管理学院)
;
Rotman School of Management, University of Toronto(多伦多大学罗特曼管理学院)
;
International Business School, Zhejiang University(浙江大学国际联合商学院)
Vision-Language Models vs Human: Perceptual Image Quality Assessment
视觉-语言模型与人类:感知图像质量评估
Imran Mehmood, Imad Ali Shah, Ming Ronnier Luo, Brian Deegan
机构
*
School of Engineering, University of Galway(Galway大学工程学院)
;
State Key Laboratory of Extreme Photonics and Instrumentation, Zhejiang University(浙江大学极端光信息获取国家重点实验室)
机构
*
Department of Computer Science and Technology, Tsinghua University, China(计算机科学与技术系,清华大学,中国)
;
College of Computer Science, Zhejiang University, China(浙江大学计算机科学学院,中国)
;
College of Computer Science, Beijing University of Posts and Telecommunications, China(北京邮电大学计算机科学学院,中国)
AI总结
本文提出连续GUI代理任务,通过引入GUI-Anchoring in Flux框架,解决GUI分布变化时持续学习稳定性问题,实验显示其优于现有方法。