ProAPO: Progressively Automatic Prompt Optimization for Visual Classification
ProAPO:逐步自动提示优化用于视觉分类
Xiangyan Qu, Gaopeng Gou, Jiamin Zhuang, Jing Yu, Kun Song, Qihao Wang, Yili Li, Gang Xiong
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
School of Information Engineering, Minzu University of China(民族大学信息工程学院)
;
University of Science and Technology Beijing(北京科技大学)
POP: Online Structural Pruning Enables Efficient Inference of Large Foundation Models
POP:在线结构剪枝实现大基础模型的高效推理
Yi Chen, Wonjin Shin, Shuhong Liu, Tho Mai, Jeongmo Lee, Chuanbo Hua, Kun Wang, Jun Liu, Joo-Young Kim
机构
*
Korea Advanced Institute of Science and Technology(韩国科学技术院)
;
University of Tokyo, Tokyo, Japan(东京大学)
;
Tokyo Institute of Technology, Tokyo, Japan(东京技术大学)
Semantically Guided Dynamic Visual Prototype Refinement for Compositional Zero-Shot Learning
语义引导的动态视觉原型精炼用于组合零样本学习
Zhong Peng, Yishi Xu, Gerong Wang, Wenchao Chen, Bo Chen, Jing Zhang, Hongwei Liu
机构
*
National Key Laboratory of Radar Signal Processing, Xidian University(雷达信号处理国家级重点实验室,西安电子科技大学)
;
Research Institute of Systems Engineering, Academy of Military Science(系统工程研究所,军事科学院)
HalluHard: A Hard Multi-Turn Hallucination Benchmark
HalluHard: 一种具有挑战性的多轮 hallucination 评估基准
Dongyang Fan, Sebastien Delsad, Nicolas Flammarion, Maksym Andriushchenko
机构
*
EPFL(苏黎世联邦理工学院)
;
ELLIS Institute Tübingen(图宾根ELLIS研究所)
;
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
;
Tübingen AI Center(图宾根人工智能中心)
机构
*
School of Electronic Engineering, Xidian University, Xi'an, China(西安电子科技大学电子工程学院)
;
School of Computer Science and Technology, Xidian University, Xi'an, China(西安电子科技大学计算机科学与技术学院)
;
College of Computer and Information, Hohai University, Nanjing, China(河海大学计算机与信息学院)
;
Institute for Infocomm Research, A*STAR, Singapore(新加坡资讯研究院)
Hallucination-Resistant Relation Extraction via Dependency-Aware Sentence Simplification and Two-tiered Hierarchical Refinement
通过依赖感知句子简化和双层级层次精炼实现抗幻觉的关系抽取
Yupei Yang, Fan Feng, Lin Yang, Wanxi Deng, Lin Qu, Biwei Huang, Shikui Tu, Lei Xu
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
University of California San Diego(加州大学圣地亚哥分校)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Alibaba Group(阿里巴巴集团)
When Ads Become Profiles: Uncovering the Invisible Risk of Web Advertising at Scale with LLMs
当广告成为资料:利用LLMs揭示大规模网络广告中的隐形风险
Baiyu Chen, Benjamin Tag, Hao Xue, Daniel Angus, Flora Salim
机构
*
The University of New South Wales(新南威尔士大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Queensland University of Technology(昆士兰理工大学)
专题命中
幻觉与鲁棒性
:multimodal large language model(abstract);分类 cs.AI
Integrating Multi-Modal Sensors: A Review of Fusion Techniques for Intelligent Vehicles
多模态传感器整合:智能车辆融合技术综述
Chuheng Wei, Ziye Qin, Ziyan Zhang, Guoyuan Wu, Matthew J. Barth
机构
*
College of Engineering, Center for Environmental Research and Technology, University of California at Riverside(工程学院、环境研究与技术中心、加州大学河滨分校)
;
School of Transportation and Logistics, Southwest Jiaotong University(交通运输与物流学院、西南交通大学)
CommentsReplacement version -- includes link to BD&S journal publication (significantly revised) in abstract, but the manuscript here remains unchanged from the original arXiv version
Training-Free and Interpretable Hateful Video Detection via Multi-stage Adversarial Reasoning
无需训练的多阶段对抗推理 hateful 视频检测
Shuonan Yang, Yuchen Zhang, Zeyu Fu
机构
*
Multimodal Intelligence Lab, Department of Computer Science, University of Exeter, United Kingdom(埃克塞特大学计算机科学系多模态智能实验室)
;
Institute for Analytics and Data Science, University of Essex, United Kingdom(埃塞克斯大学分析与数据科学研究所)
CommentsAccepted at ICASSP 2026. \c{opyright} 2026 IEEE. This is the author accepted manuscript. The final published version will be available via IEEE Xplore
Song Tang, Wenxin Su, Mao Ye, Boyu Wang, Xiatian Zhu
机构
*
Institute of Machine Intelligence, University of Shanghai for Science and Technology(上海理工大学机器智能研究院)
;
TAMS Group, Department of Informatics, Universität Hamburg(汉堡大学信息学院TAMS小组)
;
School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院)
UNCAP: Uncertainty-Guided Neurosymbolic Planning Using Natural Language Communication for Cooperative Autonomous Vehicles
UNCAP:基于自然语言通信的不确定性引导神经符号规划
Neel P. Bhatt, Po-han Li, Kushagra Gupta, Rohan Siva, Daniel Milan, Alexander T. Hogue, Sandeep P. Chinchali, David Fridovich-Keil, Zhangyang Wang, Ufuk Topcu
机构
*
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
RAL2M: Retrieval Augmented Learning-To-Match Against Hallucination in Compliance-Guaranteed Service Systems
RAL2M:检索增强的学习-匹配以对抗幻觉在合规保障服务系统中
Mengze Hong, Di Jiang, Jiangtao Wen, Zhiyang Su, Yawen Li, Yanjie Sun, Guan Wang, Chen Jason Zhang
机构
*
Hong Kong Polytechnic University(香港理工大学)
;
New York University Shanghai(纽约大学上海分校)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)