Benchmarking MRI Representations for Deep Learning-Based Focal Cortical Dysplasia Segmentation
基于深度学习的局灶性皮质发育异常分割的MRI表征基准测试
Soumen Ghosh, John Phamnguyen, Amit Soni Arya, Subhojit Mandal, Tilottama Goswami, Rajat Vashistha
机构
*
The University of Queensland(昆士兰大学)
;
I-MED Radiology(I-MED放射学)
;
Bennett University(贝内特大学)
;
Indian Institute of Technology Madras(印度马德拉斯理工学院)
;
University College of Engineering, Osmania University(奥斯曼尼亚大学工程学院)
;
Mater Hospital(玛特医院)
;
Royal Brisbane and Women’s Hospital(皇家布里斯班和妇女医院)
CHM-Net: Center Heatmap-driven Macro-Micro Modeling Network for MRI-based Microbial Density Stratification
CHM-Net:基于中心热图驱动的宏观-微观建模网络用于基于MRI的微生物密度分层
Jiaming Liang, Haolin Chen, Tingting Li, Bowen Yu, Qianyan Long, Tinghe Zhang, Xi Zhong, Xiaowei Hu, Xiaoqi Sheng, Hongmin Cai
机构
*
School of Computer Science and Engineering(计算机科学与工程学院)
;
School of Future Technology(未来技术学院)
;
Department of Medical Imaging(医学影像科)
;
Affiliated Cancer Hospital, Guangzhou Medical University(广州医学院附属癌症医院)
Capabilities of Claude Fable 5 on Biomedical Challenge Problems
Claude Fable 5在生物医学挑战问题上的能力
Dominic Okonkwo, Magnus Hodgson, Temitope I. David, Susan Adanna Ihejirika
机构
*
School of Computing, University of Georgia(佐治亚大学计算学院)
;
Department of Chemistry, University of Illinois(伊利诺伊大学化学系)
;
Institute of Bioinformatics, University of Georgia(佐治亚大学生物信息学研究所)
Andrea Boscolo Camiletto, Rishabh Dabral, Eduardo Alvarado, Thabo Beeler, Marc Habermann, Christian Theobalt
机构
*
Max Planck Institute for Informatics(马克斯·普朗克信息研究所)
;
Saarbrücken Research Center for Visual Computing, Interaction and AI(萨尔布吕肯视觉计算、交互与人工智能研究中心)
;
Google(谷歌)
Comparative Study of Domain-adapted VLMs for General Document Visual Question Answering
用于通用文档视觉问答的领域适应视觉语言模型的比较研究
Miguel Lopez-Duran, Elena Marrero, Julian Fierrez, Marta Robledo-Moreno, Ruben Vera-Rodriguez, Daniel DeAlcala, Aythami Morales, Ruben Tolosana, Oscar Delgado, Alvaro Ortigosa, Javier Ortega-Garcia
机构
*
Universidad Autónoma de Madrid (UAM)(马德里自治大学)
;
BiometricsAI(生物识别人工智能)
机构
*
Central Conservatory of Music, China(中央音乐学院)
;
Beijing Institute for General Artificial Intelligence(北京通用人工智能研究院)
;
Tianjin Conservatory of Music(天津音乐学院)
;
Beijing Electronic Science and Technology Institute(北京电子科技研究所)
;
Peking University(北京大学)
CoMind: Understanding Collaborative Human Activity from Multiple Minds and Views
CoMind:从多视角理解人类协作活动
Alexey Gavryushin, Dingxi Zhang, Zhao Huang, Alexandros Delitzas, Jiaqi Chen, Ben Ellis, Cedric Zöllner, Manthan Patel, Manuel Kaufmann, Marc Pollefeys, Xi Wang
机构
*
ETH Zurich(苏黎世联邦理工学院)
;
MPI for Informatics(马克斯·普朗克信息研究所)
;
Microsoft Switzerland(微软瑞士公司)
;
TU Munich(慕尼黑工业大学)
Benchmarking the Robustness of Autonomous Driving to Environmental Illusions: A Lane Perception Perspective
从车道感知角度评估自动驾驶对环境错觉的鲁棒性:基准测试
Tianyuan Zhang, Xianglong Liu, Aishan Liu, Lu Wang, Yitong Zhang, Peng Yue, Mingchuan Zhang, Siyuan Liang, Dacheng Tao
机构
*
SKLCCSE, the School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院软件安全技术与工程北京市重点实验室)
;
the School of Cyber Science and Technology, Sun Yat-sen University(中山大学网络空间科学与技术学院)
;
Henan University of Science and Technology(河南科技大学)
;
the School of Computing, National University of Singapore(新加坡国立大学计算学院)
;
College of Computing & Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
Ground3D-LMM: Fine-Grained 3D Point Grounding and Spatial Reasoning with LMM
Ground3D-LMM:基于LMM的细粒度3D点接地与空间推理
Amol Harsh, Zongyan Han, Jean Lahoud, Ye Liu, Rao Muhammad Anwer, Hisham Cholakkal, Salman Khan, Fahad Khan
机构
*
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Linköping University(林雪平大学)
TacReasoner: A Dynamic Tactile-Language Framework for Interactive Reasoning in Real-World Scenarios
TacReasoner:面向真实场景交互推理的动态触觉-语言框架
Kailin Lyu, Di Wu, Long Xiao, Jianning Zeng, Jianwei He, Chang Lin, Lianyu Hu, Lin Shu, Jie Hao, Ce Hao
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Beijing Zhongguancun Academy(北京中关村科学城)
;
Nanyang Technological University(南洋理工大学)
;
Guangdong Institute of Artificial Intelligence and Advanced Computing(广东省人工智能与先进计算研究院)
Autonomous Information Seeking: A Roadmap for Agentic Recommender Systems
自主信息寻求:智能推荐系统路线图
Xinyu Lin, Yashar Deldjoo, Sunhao Dai, Honghui Bao, Xiaopeng Ye, Fatemeh Nazary, Wenjie Wang, Tommaso Di Noia, Jun Xu, Tat-Seng Chua
机构
*
National University of Singapore(新加坡国立大学)
;
Polytechnic University of Bari(巴里理工大学)
;
Renmin University of China(中国人民大学)
;
University of Science and Technology of China(中国科学技术大学)
Agentic and Generative AI for Open-Source Intelligence and Cyber Investigations: Taxonomy, Evaluation, Challenges, and Future Directions
用于开源情报和网络调查的智能与生成式人工智能:分类法、评估、挑战及未来方向
Eduardo Almeida Palmieri, Mohamed Chahine Ghanem, Dipo Dunsin, Zubair Baig, Ed de Quincey, Kim-Kwang Raymond Choo
机构
*
School of Computer Science and Mathematics, Keele University(基尔大学计算机科学与数学学院)
;
Cybersecurity Institute, University of Liverpool(利物浦大学网络安全研究所)
;
Department of Applied Computing IICL, University of Wales Trinity Saint David(威尔士特里尼达大学应用计算系)
;
Deakin Cyber Research and Innovation Hub, Deakin University(德金大学网络安全研究与创新中心)
;
Department of Information Systems and Cyber Security, The University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校信息系与网络安全系)
Interpretable machine learning predicts Parkinson's disease severity using motion-corrected QSM MRI and multiband multiecho fMRI features
可解释机器学习利用运动校正QSM MRI和多波段多回波fMRI特征预测帕金森病严重程度
Aixa X. Andrade
机构
*
Lyda Hill Department of Bioinformatics(Lyda Hill 生物信息学系)
;
Department of Biomedical Engineering(生物医学工程系)
;
University of Texas Southwestern Medical Center, Dallas, Texas, USA(德克萨斯西南医学中心,德克萨斯州达拉斯)
A harmonised dataset for Earth system foundation models
用于地球系统基础模型的统一数据集
Carlos Rodriguez-Pardo, Massimo Tavoni
机构
*
Politecnico di Milano, Department of Management, Economics and Industrial Engineering(米兰理工大学管理、经济与工业工程系)
;
RFF-CMCC European Institute on Economics and the Environment(资源未来研究所-欧洲经济与环境CMCC研究所)
;
CMCC Foundation - Euro-Mediterranean Center on Climate Change(CMCC基金会-欧洲地中海气候变化中心)
Evaluation of Medical Vision Language Models HuluMed and MedGemma, and general purpose chatbots Gemma 3, ChatGPT Plus, and Claude Pro on real previously unseen wound images
医学视觉语言模型 HuluMed 和 MedGemma 以及通用聊天机器人 Gemma 3、ChatGPT Plus 和 Claude Pro 在真实未见伤口图像上的评估
Yunzhe Xue, Mohammed Saim Ahmed Quadri, Neal Panse, Justin W. Ady, Usman Roshan
机构
*
Department of Computer Science, New Jersey Institute of Technology(新泽西理工学院计算机科学系)
;
Vascular and Endovascular Surgery, Robert Wood Johnson Hospital(罗伯特·伍德·约翰逊医院血管外科)
;
Department of Data Science, New Jersey Institute of Technology(新泽西理工学院数据科学系)
专题命中
多模态评测
:multimodal(abstract);分类 cs.CV
AI总结
本研究评估了六种视觉语言模型在慢性伤口分析任务上的表现,发现通用模型 ChatGPT 和 Claude 显著优于医学专用模型,表明广泛的多模态推理能力比领域知识更重要。