IndusAgent: Reinforcing Open-Vocabulary Industrial Anomaly Detection with Agentic Tools
IndusAgent: 通过智能工具增强开放词汇工业异常检测
Rongbin Tan, Fangfang Lin, Zhenlong Yuan, Min Qiu, Kejin Cui, Mengmeng Wang, Yi Wang, Zijian Song, Zhiyuan Wang, Jiyuan Wang, Yue Wang, Shuhan Song§, Huawei Cao
机构
*
State Key Lab of Processors, Institute of Computing Technology, CAS(处理器国家重点实验室,计算技术研究所,中国科学院)
;
Santa Clara University(圣克拉拉大学)
;
LongCat Team(LongCat团队)
;
Independent Researcher(独立研究者)
;
New York University(纽约大学)
;
Sun Yat-sen University(孙中山大学)
;
Nanyang Technological University(南洋理工大学)
;
Stanford University(斯坦福大学)
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)
机构
*
Dream-X Team(Dream-X团队)
;
Stanford University(斯坦福大学)
;
National University of Singapore(新加坡国立大学)
;
Independent Researcher(独立研究者)
;
Case Western Reserve University(凯斯西储大学)
;
UC Santa Cruz(加州大学圣克鲁兹分校)
CommentsAccepted at the IMAGE'25 Workshop (PCW-11), Society of Exploration Geophysicists (SEG). Published version available at https://doi.org/10.1190/image2025-w11-03.1
LLM-Powered Flood Depth Estimation from Social Media Imagery: A Vision-Language Model Framework with Mechanistic Interpretability for Transportation Resilience
机构
*
Wangxuan Institute of Computer Technology, Peking University(北京大学王学苑计算机技术研究所)
;
State Key Laboratory of General Artificial Intelligence(通用人工智能国家重点实验室)
CommentsPlease cite as: W. Li, S. Manickam, Y. -W. Chong and S. Karuppayah, "PhishDebate: An LLM-Based Multi-Agent Framework for Phishing Website Detection," 2025 IEEE International Conference on Big Data (BigData), Macau, China, 2025, pp. 6606-6615, doi: 10.1109/BigData66926.2025.11401440
Journal ref2025 IEEE International Conference on Big Data (BigData), Macau, China, 2025, pp. 6606-6615
Factuality Matters: When Image Generation and Editing Meet Structured Visuals
事实性至关重要:当图像生成与编辑遇见结构化视觉
Le Zhuo, Songhao Han, Yuandong Pu, Boxiang Qiu, Sayak Paul, Yue Liao, Yihao Liu, Jie Shao, Xi Chen, Si Liu, Hongsheng Li
机构
*
CUHK MMLab(香港大学多模态实验室)
;
Beihang University(北京航空航天大学)
;
Krea AI(Krea人工智能)
;
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai AI Lab(上海人工智能实验室)
;
Hugging Face
;
National University of Singapore(新加坡国立大学)
;
ByteDance(字节跳动)
;
The University of Hong Kong(香港大学)
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Nanjing University(南京大学)
;
Wuhan University(武汉大学)
机构
*
School of Biomedical Engineering, Southern Medical University(生物医学工程学院,南方医科大学)
;
School of Biomedical Engineering, Shanghai Jiaotong University(生物医学工程学院,上海交通大学)
;
Department of Electronic Engineering, Chinese University of Hong Kong(电子工程系,中国香港大学)
;
Faculty of Dentistry, The University of Hong Kong(牙科学院,香港大学)
;
Department of Nuclear Medicine, The Second Affiliated Hospital of Guangzhou University of Chinese Medicine(核医学科,广州中医药大学第二附属医院)
;
PET Center, Department of Nuclear Medicine, Guangdong Provincial People’s Hospital, Southern Medical University(PET中心,核医学科,广东省人民医院,南方医科大学)
;
Department of Nuclear Medicine, Nanfang Hospital, Southern Medical University(核医学科,南芳医院,南方医科大学)
;
Division of Nuclear Medicine and Molecular Imaging, Geneva University Hospitals(核医学与分子影像学部,日内瓦大学医院)
;
Departments of Radiology, Physics, and Biomedical Engineering, The University of British Columbia(放射学、物理和生物医学工程系,不列颠哥伦比亚大学)
;
Medical Artificial Intelligence Laboratory, Westlake University(医学人工智能实验室,西湖大学)
DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models
DVGBench: 面向无人机影像的隐式到显式视觉 grounding 评估基准
Yue Zhou, Jue Chen, Zilun Zhang, Penghui Huang, Ran Ding, Zhentao Zou, PengFei Gao, Yuchen Wei, Ke Li, Xue Yang, Xue Jiang, Hongxin Yang, Jonathan Li
机构
*
Hinton STAI Institute(Hinton STAI研究所)
;
Key Laboratory of Geographic Information Science (Ministry of Education), East China Normal University(地理信息科学重点实验室(教育部))
;
School of Geospatial Artificial Intelligence, East China Normal University(地理空间人工智能学院)
;
Zhejiang University(浙江大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Information Engineering University(信息工程大学)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Fudan University(复旦大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
KiVA: Kid-inspired Visual Analogies for Testing Large Multimodal Models
KiVA:儿童启发的视觉类比用于测试大多模态模型
Eunice Yiu, Maan Qraitem, Anisa Noor Majhi, Charlie Wong, Yutong Bai, Shiry Ginosar, Alison Gopnik, Kate Saenko
机构
*
University of California, Berkeley(加州大学伯克利分校)
;
Boston University(波士顿大学)
;
Google DeepMind(谷歌DeepMind)
;
Toyota Technological Institute at Chicago(芝加哥丰田技术研究所)