TD-VAD: Breaking Visual Dependence in Video Anomaly Detection with Text-Driven Learning
TD-VAD:通过文本驱动学习打破视频异常检测中的视觉依赖
机构 * School of Intelligence Science and Technology, Nanjing University(南京大学智能科学与技术学院) ; School of Artificial Intelligence Engineering, Hefei Institute of Technology(合肥工业大学人工智能工程学院) ; China Mobile Zijin Innovation Institute(中国移动紫金创新研究院) ; School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院)
AI总结 针对现有视频异常检测依赖视觉数据的问题,提出TD-VAD方法,利用LLM生成的文本描述训练模型,结合事件演化因果注意力模块与CLIP编码器,在XD-Violence和UCF-Crime上性能大幅优于现有方法。
Comments Accepted to ICML2026