Comments9 pages. v2: results updated to July 2026 leaderboard (17 models). Accepted at the 2nd Workshop on Knowledge-Intensive Multimodal Reasoning (KnowledgeMR) at CVPR 2026 (non-archival), under the former title "PDFParse: A Benchmark for Grounded Multimodal Reasoning over Professional PDF Documents". Dataset: https://huggingface.co/datasets/surgeai/GDP.pdf ; Code: https://github.com/surge-ai/gdp-pdf
机构
*
School of Information Science and Technology, University of Science and Technology of China(科学技术大学信息科学与技术学院)
;
ByteDance Intelligent Creation(字节跳动智能创作)
;
School of Computer Science and Technology, Harbin Institute of Technology (Weihai)(哈尔滨工业大学(威海)计算机科学与技术学院)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院)
StructuredEdit: Constraint-Aware Graphic Design Editing via Differentiable Parameter Propagation
StructuredEdit:通过可微参数传播实现约束感知的平面设计编辑
Veeramanohar Avudaiappan, Ritwik Murali
机构
*
Department of Electrical and Electronics Engineering, Amrita School of Engineering, Coimbatore, Amrita Vishwa Vidyapeetham India(电子与电子工程系,阿米特拉工程学院,科伊巴托尔,阿米特拉世界学院,印度)
;
Department of Computer Science and Engineering, Amrita School of Computing, Coimbatore, Amrita Vishwa Vidyapeetham India(计算机科学与工程系,阿米特拉计算学院,科伊巴托尔,阿米特拉世界学院,印度)
CommentsAccepted to the 43rd International Conference on Machine Learning (ICML 2026). 22 pages, 11 figures. Code and dataset available at https://github.com/zimoqingfeng/MORE
机构
*
Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)
;
SparcAI Inc(SparcAI公司)
;
University of Science and Technology of China(中国科学技术大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Nanyang Technological University(南洋理工大学)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
Advancing WordArt-Oriented Scene Text Recognition: Datasets and Methods
推进面向艺术字的场景文本识别:数据集与方法
Xingsong Ye, Yongkun Du, Jiaxin Zhang, Haojie Zhang, Chong Sun, Chen Li, Jing Lyu, Zhineng Chen
机构
*
Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究所)
;
Shanghai Key Laboratory of Multimodal Embodied AI, Fudan University(复旦大学上海市多模态具身人工智能重点实验室)
;
WeChat Vision, Tencent Inc.(腾讯微信视觉团队)
;
South China University of Technology(华南理工大学)
CommentsPublished in the Proceedings of the 51st Euromicro Conference on Software Engineering and Advanced Applications, SEAA 2025. Lecture Notes in Computer Science, volume 16082, pages 143-158. Springer, 2026
PlotPick: AI-powered batch extraction of numerical data from scientific figures
PlotPick: 基于AI的科学图表中数值数据批量提取工具
Tommy Carstensen
机构
*
Copenhagen Research Centre for Biological and Precision Psychiatry(哥本哈根生物与精准精神病学研究中心)
;
Mental Health Centre Copenhagen(哥本哈根心理健康中心)
;
Copenhagen University Hospital(哥本哈根大学医院)