Separating Clicks from Baits: Using Large Language Models to Detect Misleading YouTube Thumbnails
从诱饵中分离点击:使用大语言模型检测误导性的YouTube缩略图
专题命中 视觉定位与Grounding :vision-language model(abstract);LLaVA(abstract)
AI总结 研究针对YouTube等平台误导性缩略图问题,提出用大语言模型的多模态检测管道,构建数据集并评估多个模型,发现Claude 3.5 Sonnet性能突出,通过失败分析为视频平台发展提供方向。
Comments Accepted to the 21st International AAAI Conference on Web and Social Media (ICWSM 2027)
Journal ref Proceedings of the International AAAI Conference on Web and Social Media, 2027