Technical Report on the CVPR 2026@AdvML Workshop Challenge
关于CVPR 2026@AdvML研讨会挑战赛的技术报告
Tianyuan Zhang, Zonglei Jing, Jiangfan Liu, Ligong Zhang, Ke Ma, Chengzhi Sun, Xiaohai Xu, Zhirui Zhang, Qianqian Xu, Qingming Huang, Hanyu Fang, Junhua Liu, Zheng Wang, Xiaoliang Liu, Yuanbo Li, Shuai Gui, Bin Wang, Menghe Zheng, Jing Nie, Hanyang Meng, Zeyang Zhang, Xiang Zhang, Yongxuan Zhu, Rui Ding, Hainan Li, Yongkang Zhang, Zhilei Zhu, Xianglong Kong, Jin Hu, Zonghao Ying, Yisong Xiao, Lei Chen, Haotong Qin, Jiakai Wang, Aishan Liu, Ruikai Li, Julia Karbing, Yinpeng Dong, Zhenfei Yin, Shao Jing, Xia Hu, Jingyi Xu, Juntao Dai, Xinyun Chen, Vishal M. Patel, Xianglong Liu, Dawn Song, Alan Yuille, Philip H. S. Torr, Dacheng Tao
机构
*
Beihang University(北京航空航天大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Tongji University(同济大学)
;
iFLYTEK Co., Ltd.(科大讯飞股份有限公司)
;
Anhui Laboratory for Safe Artificial Intelligence in the Yangtze River Delta(长三角安全人工智能安徽实验室)
;
Wenzhou Business College(温州商学院)
;
Jiangnan University(江南大学)
;
Guangzhou City University of Technology(广州理工学院)
;
Inceptio Technology(智元机器)
;
Institute of Dataspace(数据空间研究所)
;
Zhongguancun Laboratory(中关村实验室)
;
Tsinghua University(清华大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
University of Oxford(牛津大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
BAAI(北京智源人工智能研究院)
;
Meta
;
Johns Hopkins University(约翰·霍普金斯大学)
;
University of California, Berkeley(加州大学伯克利分校)
;
Nanyang Technological University(南洋理工大学)
Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning
通过深度原生结构推理实现准确、跨学科和透明的结构-属性理解
Chen Tang, Yizhou Wang, Jianyu Wu, Lintao Wang, Shixiang Tang, Pengze Li, Encheng Su, Jun Yao, Jiabei Xiao, Yuqi Shi, Jielan Li, Hongxia Hao, Zhangyang Gao, Fang Wu, Ben Fei, Xiangyu Yue, Pan Tan, Bozitao Zhong, Jinouwen Zhang, Aoran Wang, Yan Lu, Jiaheng Liu, Xinzhu Ma, Liang Hong, Mingyue Zheng, Phil Torr, Bowen Zhou, Wanli Ouyang, Lei Bai
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Fudan University(复旦大学)
;
University of Sydney(悉尼大学)
;
Nanjing University(南京大学)
;
University of Oxford(牛津大学)
;
The University of Science and Technology of China(中国科学技术大学)
;
Drug Discovery and Design Center, State Key Laboratory of Drug Research, Shanghai Institute of Materia Medica, Chinese Academy of Sciences(药物发现与设计中心、国家药物研究重点实验室、上海中医药材料医学研究所、中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Stanford University(斯坦福大学)
机构
*
School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院)
;
Institute of Plasma Physics, Chinese Academy of Sciences(中国科学院等离子体物理研究所)
;
State Key Laboratory of Opto-Electronic Information Acquisition and Protection Technology, Anhui University(安徽大学光电子信息获取与控制技术国家重点实验室)
HCSU: A Dataset and Benchmark for Fine-Grained Historical Calligraphy Style Understanding
HCSU:用于细粒度历史书法风格理解的数据集和基准测试
Yinsheng Yao, Yan Liu, Chen Ye
机构
*
School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院)
;
The Key Laboratory of Embedded System and Service Computing, Ministry of Education(教育部嵌入式系统与服务计算重点实验室)
Transition Information Density: Morphological Trajectories, Synesthetic Perception, and Structured Interpolation in Neural Training (or: The Synesthetic AI)
过渡信息密度:神经训练中的形态轨迹、联觉感知和结构化插值(或:联觉人工智能)
Sam Mao
机构
*
New York University(纽约大学)
;
Interactive Media Arts(互动媒体艺术)
Comments38 pages, 9 figures, 4 tables. Empirical results from structured interpolation training across four representational mediums. Pipeline scripts, experimental data, and the Synesthesia Grid algorithm available upon reasonable request
EduArt: An educational-level benchmark for evaluating art history knowledge in large language models
EduArt:评估大型语言模型艺术史知识的教育级基准
Gianmarco Spinaci, Lukas Klic, Giovanni Colavizza
机构
*
University of Bologna(博洛尼亚大学)
;
Villa i Tatti – The Harvard University Center for Italian Renaissance Studies(哈佛大学意大利文艺复兴研究中心(I Tatti))
;
University of Copenhagen(哥本哈根大学)
JL1-CC&QA: Extending the JL1-CD Benchmark with Change Captioning and Question Answering
JL1-CC&QA:扩展JL1-CD基准,增加变化描述和问答
Ziyuan Liu, Ruifei Zhu, Ouqiao Ma, Yuantao Gu
机构
*
Department of Electronic Engineering, Beijing National Research Center for Information Science and Technology, Tsinghua University(清华大学电子工程系,北京信息科学与技术国家研究中心)
;
Chang Guang Satellite Technology Co., Ltd. (CGSTL)(长光卫星技术股份有限公司)
;
College of Communications Engineering, Army Engineering University of PLA(中国人民解放军陆军工程大学通信工程学院)
Automated sign detection across the Electronic Babylonian Library: A large-scale dataset and end-to-end cuneiform OCR pipeline
电子巴比伦图书馆中的自动楔形文字符号检测:大规模数据集与端到端楔形文字OCR流水线
Wentao Che, Esteban Garcés Arias, Asim Niaz, Andreas Bender, Enrique Jiménez
机构
*
Institute of Assyriology and Hittite Studies, LMU Munich(慕尼黑大学亚述学与赫梯学研究所)
;
Department of Statistics, LMU Munich(慕尼黑大学统计系)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
StylisticBias: A Few Human Visual Cues Drive Most Social Biases in MLLMs
StylisticBias: 少数人类视觉线索驱动多模态大语言模型中的大部分社会偏见
Shaghayegh Kolli, Timo Cavelius, Nafiseh Nikeghbal, Samantha Dalal, Jana Diesner
机构
*
Technical University of Munich(慕尼黑工业大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
Princeton Center for Information and Technology Policy(普林斯顿信息与技术政策中心)
NRITYAM: Language Models Meet Art and Heritage of Dance
NRITYAM:语言模型遇见舞蹈的艺术与遗产
Punit Kumar Singh, Niladri Ghosh, Advait Joshiınst, Shailee Choudhary, Michael Färber, Haiqin Yang
机构
*
Shenzhen Technology University(深圳技术大学)
;
New Delhi Institute of Management(新德里管理学院)
;
Technische Universität Dresden(德累斯顿工业大学)
;
Ramakrishna Mission Vivekananda Educational and Research Institute(罗摩克里希纳传道会维韦卡南达教育与研究学院)
;
Indian Institute of Technology(印度理工学院)
;
Swami Vivekananda Institute of Technology(斯瓦米·维韦卡南达技术学院)
;
GuangDong Engineering Technology Research Center of Edge Intelligence(广东省边缘智能工程技术研究中心)
TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs
TABVERSE:大语言模型与视觉语言模型中跨格式表格理解的基准测试
Momina Ahsan, Sarfraz Ahmad, Ming Shan Hee, Roy Ka-Wei Lee, Preslav Nakov
机构
*
Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学)
;
Singapore University of Technology and Design (SUTD)(新加坡科技设计大学)
ChinaHeritaQA: A Culturally-Grounded Visual Question Answering Dataset for World Heritage Sites in China
ChinaHeritaQA:面向中国世界遗产地的文化基础视觉问答数据集
Yi Zhang, Bolei Ma, Yong Cao, Chengyan Wu, Daniel Hershcovich, Anna-Carolina Haensch
机构
*
LMU Munich(慕尼黑大学)
;
FAU Erlangen-Nuremberg(埃尔朗根-纽伦堡大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
University of Tübingen & Tübingen AI Center(图宾根大学与图宾根人工智能中心)
;
Sun Yat-sen University(中山大学)
;
University of Copenhagen(哥本哈根大学)
;
University of Maryland, College Park(马里兰大学帕克分校)