VideoGAIA: A Benchmark for General AI Assistants on Agentic Video Understanding
VideoGAIA:面向通用人工智能助手的智能体视频理解基准
Fan Zhang, Guangming Yao, Jinyang Wu, Hao Wu, Zheng Lian, Xinyu Geng, Jingdong Chen, Yi Yuan, Pheng-Ann Heng
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Ant Group(蚂蚁集团)
;
Tsinghua University(清华大学)
;
Tongji University(同济大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
AutoSchema: Live Schema Grounding for Agentic Text-to-Sparql over Heterogeneous Knowledge Graphs
AutoSchema:面向异构知识图谱的智能体文本转SPARQL的实时模式接地
Yiming Zhang, Koji Tsuda
机构
*
The University of Tokyo(东京大学)
;
National Institute for Materials Science(国立材料科学研究所)
;
RIKEN Center for Advanced Intelligence Project(理化学研究所高级智能项目中心)
PULSE: Agentic Investigation with Passive Sensing for Proactive Affective Intervention in Cancer Survivorship
PULSE:基于被动感知的代理探究用于癌症幸存者的主动干预
Zhiyuan Wang, Subigya Nepal, Ariful Islam, Indrajeet Ghosh, Xinyu Chen, Katharine E. Daniel, Laura E. Barnes, Philip Chow
机构
*
Department of Systems and Information Engineering, University of Virginia(系统与信息工程系,弗吉尼亚大学)
;
Center for Behavioral Health and Technology, University of Virginia(行为健康与技术中心,弗吉尼亚大学)
;
Department of Computer Science, University of Virginia(计算机科学系,弗吉尼亚大学)
"LLM Agent Performance" Is Not a Single Evaluation Target
基于LLM的智能体评估统一框架的必要性
Pengyu Zhu, Li Sun, Philip S. Yu, Sen Su
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
University of Illinois Chicago(伊利诺伊大学芝加哥分校)
;
Chongqing University of Posts and Telecommunications(重庆邮电大学)