arXivDaily arXiv每日学术速递 周一至周五更新

科学与医疗

医学 AI

医学智能、临床 AI、医学影像、病理、诊断和医疗健康大模型。

2026-04-17 至 2026-04-17 共收录 33 信号源:cs.CV, cs.LG, q-bio, eess.IV, eess.SP

1. 医学影像 16 篇

2604.14800 2026-04-17 eess.IV cs.CV physics.med-ph 92%

Generative Modeling of Complex-Valued Brain MRI Data

复杂值脑部MRI数据的生成建模

Marco Schlimbach, Moritz Rempe, Jessica Mnischek, Lukas T. Rotkopf, Jens Weingarten, Jens Kleesiek, Kevin Kröninger

机构 * Department of Physics, Technical University Dortmund(技术大学多特蒙德物理系) Institute for AI in Medicine (IKIM), University Hospital Essen(医学人工智能研究所(IKIM),埃森大学医院) Cancer Research Center Cologne Essen (CCCE), University Medicine Essen(科隆埃森癌症研究中心(CCCE),埃森大学医学中心) RACOON Study Group, Site Essen(RACOON研究小组,埃森站点) German Cancer Consortium (DKTK), Partner Site Essen(德国癌症联合会(DKTK),埃森合作伙伴站点) Medical Faculty and Faculty of Computer Science, University of Duisburg-Essen(杜伊斯堡-埃森大学医学系和计算机科学系) Division of Radiology, German Cancer Research Center (DKFZ), Im Neuenheimer Feld 280(放射学系,德国癌症研究中心(DKFZ),新海德费尔德280号)

专题命中 医学影像 :MRI(title,title_cn);pathology(abstract);diagnosis(abstract);分类 cs.CV、eess.IV

AI总结 本文提出一种联合建模复值MRI幅度和相位信息的生成框架,通过变分自编码器和流匹配模型生成高质量合成数据,提升异常组织检测性能。

Comments 16 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.05882 2026-04-17 eess.IV cs.CV cs.LG cs.NE 90%

Frame forecasting in cine MRI using the PCA respiratory motion model: comparing recurrent neural networks trained online and transformers

利用PCA呼吸运动模型在动态MRI中进行帧预测:比较在线训练的循环神经网络和变换器

Michel Pohl, Mitsuru Uesaka, Hiroyuki Takahashi, Kazuyuki Demachi, Ritu Bhusal Chhatkuli

机构 * The University of Tokyo(东京大学) National Institutes for Quantum and Radiological Science and Technology(量子与辐射科学和技术国家研究所)

专题命中 医学影像 :MRI(title,title_cn);分类 cs.CV、cs.LG、eess.IV

AI总结 本文研究了利用PCA呼吸运动模型在动态MRI中进行帧预测,比较了在线训练的循环神经网络和变换器在处理呼吸运动补偿中的性能差异。

Comments 43 pages, 19 figures. Revised version with minor corrections and improved figures and language. Accepted for publication in Computerized Medical Imaging and Graphics

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04588 2026-04-17 cs.CV 87%

3D Conditional Image Synthesis of Left Atrial LGE MRI from Composite Semantic Masks

基于复合语义掩码的左心房LGE MRI的3D条件图像合成

Yusri Al-Sanaani, Rebecca Thornhill, Sreeraman Rajan

机构 * Systems and Computer Engineering, Carleton University(卡尔顿大学系统与计算机工程系) Department of Radiology, University of Ottawa(渥太华大学放射科部)

专题命中 医学影像 :MRI(title,title_cn);分类 cs.CV

AI总结 本文提出利用3D条件生成模型增强稀少LGE训练数据,提升左心房分割性能,SPADE-LDM生成最逼真图像,使Dice分数显著提升。

Comments This work has been published in the Proceedings of the 2025 IEEE International Conference on Imaging Systems and Techniques (IST). The final published version is available via IEEE Xplore

Journal ref 2025 IEEE International Conference on Imaging Systems and Techniques (IST)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06591 2026-04-17 cs.CV 87%

Hybrid Swin Attention Networks for Simultaneously Low-Dose PET and CT Denoising

混合Swin注意力网络用于同时低剂量PET和CT去噪

Yichao Liu, Hengzhi Xue, YueYang Teng, Junwen Guo

机构 * organization= IWR, Heidelberg University , city= Heidelberg , postcode= 69120 , state= Baden Württemberg , country= Germany organization= College of Medicine Biological Information Engineering, Northeastern University , city= Shenyang , postcode= 110169 , state= Liaoning , country= China organization= Key Laboratory of Intelligent Computing in Medical Image, Ministry of Education , city= Shenyang , postcode= 110169 , state= Liaoning , country= China organization= Department of Epidemiology \& Global Health, Umeå University , addressline= , city= Umeå , postcode= 90187 , country= Sweden

专题命中 医学影像 :CT(title,title_cn);分类 cs.CV

AI总结 本文提出混合Swin注意力网络HSANet,结合高效全局注意力模块和混合上采样模块,提升低剂量PET和CT去噪性能,同时保持模型轻量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11176 2026-04-17 cs.CV 84%

Precision Synthesis of Multi-Tracer PET via VLM-Modulated Rectified Flow for Stratifying Mild Cognitive Impairment

多示踪PET的高精度合成通过VLM调制的校正流用于区分轻度认知障碍

Tuo Liu, Shuijin Lin, Shaozhen Yan, Haifeng Wang, Jie Lu, Jianhua Ma, Chunfeng Lian

机构 * School of Mathematics and Statistics, Xi'an Jiaotong University(西安交通大学数学与统计学学院) Key Laboratory of Biomedical Information Engineering of Ministry of Education, School of Life Science and Technology, Xi'an Jiaotong University(教育部生物医学信息工程重点实验室,西安交通大学生命科学与技术学院) Department of Radiology and Nuclear Medicine, Xuanwu Hospital, Capital Medical University(首都医科大学宣武医院放射科与核医学科) Research Center for Intelligent Medical Equipment and Devices (IMED), Xi'an Jiaotong University(智能医疗设备与器件研究中心(IMED),西安交通大学)

专题命中 医学影像 :MRI(summary_cn,abstract);diagnosis(abstract);分类 cs.CV

AI总结 本文提出DIReCT$++$模型,结合MRI和临床信息,通过校正流和视觉语言模型生成高保真多示踪PET图像,实现轻度认知障碍的精准分层。

Comments Added supplementary material

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14216 2026-04-17 cs.MM cs.AI cs.CL cs.CV cs.GR cs.LG 82%

Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis

Neuro-Oracle:一种具有轨迹意识的代理RAG框架用于可解释性癫痫手术预后

Aizierjiang Aiersilan, Mohamad Koubeissi

机构 * The George Washington University(乔治华盛顿大学)

专题命中 医学影像 :MRI(summary_cn,abstract);分类 cs.CV、cs.LG

AI总结 本文提出Neuro-Oracle框架,通过三维Siamese对比编码器提取术前术后MRI变化,检索历史相似手术轨迹,并利用量化Llama-3-8B推理代理生成自然语言预后,验证了轨迹意识检索架构的可行性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14545 2026-04-17 cs.RO 82%

CT-VIR: Continuous-Time Visual-Inertial-Ranging Fusion for Indoor Localization with Sparse Anchors

CT-VIR:基于稀疏锚点的连续时间视觉-惯性-测距融合用于室内定位

Yu-An Liu, Li Zhang

机构 * School of Mathematics, Hefei University of Technology(合肥工业大学数学学院)

专题命中 医学影像 :CT(title,title_cn)

AI总结 本文提出一种基于样条的连续时间状态估计方法,用于视觉-惯性-测距融合定位,通过预处理阶段构建虚拟锚点并拒绝异常值,提升几何退化和测距可靠性,采用B样条参数化姿态轨迹并联合优化控制点和辅助参数,实验证明其有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.15059 2026-04-17 cs.CV 81%

Attention-Gated Convolutional Networks for Scanner-Agnostic Quality Assessment

注意力门控卷积网络用于无扫描仪质量评估

Chinmay Bakhale, Anil Sao

机构 * Indian Institute of Technology, Bhilai, India(印度比哈尔理工学院)

专题命中 医学影像 :MRI(summary_cn,abstract);分类 cs.CV

AI总结 本文提出一种混合CNN-注意力框架,用于鲁棒且跨站点的MRI质量评估,通过局部空间特征提取和多头交叉注意力机制,实现运动伪影的优先识别与噪声过滤,验证了在不同扫描仪环境下的泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14849 2026-04-17 cs.CV cs.AI 79%

Efficient Search of Implantable Adaptive Cells for Medical Image Segmentation

高效搜索可植入自适应细胞用于医学图像分割

Emil Benedykciuk, Marcin Denkowski, Grzegorz M. Wójcik

机构 * Institute of Computer Science and Mathematics, Maria Curie Sklodowska University(马里亚·科洛多夫斯卡大学计算机科学与数学研究所)

专题命中 医学影像 :medical image(title,abstract);分类 cs.CV

AI总结 本文提出IAC-LTH框架,通过追踪操作重要性分布,加速搜索过程,减少计算成本,提升分割性能。

Comments 20 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14506 2026-04-17 cs.CV 79%

Co-distilled attention guided masked image modeling with noisy teacher for self-supervised learning on medical images

协同蒸馏引导的掩码图像建模与噪声教师用于医学图像的自监督学习

Jue Jiang, Aneesh Rangnekar, Harini Veeraraghavan

机构 * Memorial Sloan Kettering Cancer Center(纪念斯隆凯特林癌症中心)

专题命中 医学影像 :medical image(title,abstract);分类 cs.CV

AI总结 本文提出DAGMaN框架,通过引导注意力的掩码机制和噪声教师提升医学图像自监督学习的效果,减少信息泄露并增强预训练难度,应用于肺结节分类、免疫治疗预测等任务。

Comments Accepted at MIDL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20218 2026-04-17 eess.IV q-bio.QM 71%

Robust Glioblastoma Segmentation and Volumetry Without T2-FLAIR: External Validation of Targeted Dropout Training

鲁棒性胶质瘤分割与体积测量无需T2-FLAIR:针对dropout训练的外部验证

Marco Öchsner, Lena Kaiser, Robert Stahl, Nathalie L. Albert, Thomas Liebig, Robert Forbrig, Jonas Reis

专题命中 医学影像 :MRI(abstract,abstract_cn);分类 q-bio、eess.IV

AI总结 本文通过外部验证证明了在无T2-FLAIR情况下,利用目标dropout训练的3D nnU-Net模型可保持胶质瘤分割性能,并显著降低全肿瘤分割误差和体积偏差。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24992 2026-04-17 cs.CV 70%

C2W-Tune: Cavity-to -Wall Transfer Learning for Thin Atrial Wall Segmentation in 3D Late Gadolinium-enhanced Magnetic Resonance

C2W-Tune:用于3D晚期钆增强磁共振中薄心房壁分割的腔体到壁迁移学习

Yusri Al-Sanaani, Rebecca Thornhill, Sreeraman Rajan

机构 * Systems and Computer Engineering, Carleton University(卡莱顿大学系统与计算机工程系) Department of Radiology, University of Ottawa(渥太华大学放射科部)

专题命中 医学影像 :MRI(abstract,abstract_cn);分类 cs.CV

AI总结 本文提出C2W-Tune框架,通过高精度左心房腔体模型提升薄壁分割精度,实验显示在2018年左心房分割挑战数据集上,Dice系数和边界误差显著降低。

Comments Submitted this to the International Conference on Artificial Intelligence in Medicine (AIME 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14316 2026-04-17 cs.AI 67%

Seeing Through Experts Eyes A Foundational Vision Language Model Trained on Radiologists Gaze and Reasoning

通过专家之眼:一个基于放射学家目光和推理训练的基础视觉语言模型

Kinhei Lee, Peiyuan Jing, Zhenxuan Zhang, Yue Yang, Tao Wang, Dominic C Marshall, Yingying Fang, Guang Yang

机构 * Bioengineering Department and Imperial-X(生物工程系和Imperial-X) Imperial College London(帝国理工学院伦敦分校) College of physics and information engineering, Fuzhou University(福州大学物理与信息工程学院) Department of Surgery and Cancer, Imperial College London(外科与癌症部门,帝国理工学院伦敦分校) National Heart and Lung Institute, Imperial College London(国家心脏和肺研究所,帝国理工学院伦敦分校) Cardiovascular Research Centre, Royal Brompton Hospital(心血管研究中心,皇家布里顿医院) School of Biomedical Engineering & Imaging Sciences, King’s College London(生物医学工程与成像科学学院,国王学院伦敦分校)

专题命中 医学影像 :medical image(abstract);radiology(abstract)

AI总结 本文提出GazeX模型,通过放射学家的眼动数据模拟专家推理过程,提升医学影像分析的准确性和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14844 2026-04-17 eess.IV cs.CV cs.LG 67%

Improving Prostate Gland Segmentation Using Transformer based Architectures

利用基于变换器的架构改进前列腺分割

Shatha Abudalou Yasin Yilmaz Yoganand Balagurunathan

机构 * Department of Machine Learning(机器学习系) Diagnostic Radiology(诊断放射科) H. Lee Moffitt Cancer Center and Research Institute(H. Lee Moffitt癌症中心和研究所)

专题命中 医学影像 :MRI(abstract);分类 cs.CV、cs.LG、eess.IV

AI总结 本文研究了变换器模型在应对读者差异和跨站点域偏移时的分割性能,通过比较UNETR、SwinUNETR与传统3D UNet,证明SwinUNETR在Dice相似度上提升显著,具备更高的鲁棒性和计算效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14263 2026-04-17 q-bio.TO cs.CV cs.LG 65%

A deep learning framework for glomeruli segmentation with boundary attention

用于肾小球分割的深度学习框架与边界注意力

Behnaz Elhaminia, Catherine King, Jiaqi Lv, Lorraine Harper, Paul Moss, Owen Cain, Dimitrios Chanouzas, Shan E Ahmed Raza

机构 * Tissue Image Analytics (TIA) Centre, Dept. of Computer Science, University of Warwick, UK(沃里克大学计算机科学系组织图像分析中心) Dept. of Immunology and Immunotherapy, University of Birmingham, UK(伯明翰大学免疫学与免疫治疗系) Renal Unit, Queen Elizabeth Hospital Birmingham, UHB NHS Foundation Trust, UK(伯明翰女王医院肾病科,UHB国家健康服务基金会信托) Dept. of Cellular Pathology, Queen Elizabeth Hospital Birmingham, UK(伯明翰女王医院细胞病理学系) School of Applied Health Sciences, University of Birmingham, UK(伯明翰大学应用健康科学学院)

专题命中 医学影像 :pathology(abstract);分类 cs.CV、cs.LG、q-bio

AI总结 本文提出一种强调边界分离的深度学习框架,通过路径学基础模型和专门设计的注意力解码器提升肾小球实例分割性能,实验表明在Dice分数和交并比上优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14720 2026-04-17 cs.CV 57%

Data Synthesis Improves 3D Myotube Instance Segmentation

数据合成提升3D肌管实例分割

David Exler, Nils Friederich, Martin Krüger, John Jbeily, Mario Vitacolonna, Rüdiger Rudolf, Ralf Mikut, Markus Reischl

机构 * Institute for Automation and Applied Informatics, Karlsruhe Institute of Technology(自动化与应用信息学院,卡尔斯鲁厄理工学院) Institute of Biological and Chemical Systems, Karlsruhe Institute of Technology(生物与化学系统研究所,卡尔斯鲁厄理工学院) CeMOS Research and Transfer Center, Technische Hochschule Mannheim(CeMOS研究与转移中心,曼海姆技术大学)

专题命中 医学影像 :biomedical(abstract);分类 cs.CV

AI总结 本文提出基于几何驱动的数据合成方法,通过多项式中心线、局部变化半径、分支结构和椭球形端帽等模型生成肌管数据,训练紧凑的3D U-Net模型,在真实数据上达到更高的IPQ指标,验证了生物物理驱动合成在标注稀缺的生物医学领域中的有效性。

Comments 4 pages, 4 figures, submitted to BMT (VDE) 2026 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 临床大模型 2 篇

2602.07529 2026-04-17 cs.LG 57%

MedVerse: Efficient and Reliable Medical Reasoning via DAG-Structured Parallel Execution

MedVerse: 通过DAG结构并行执行实现高效可靠的医学推理

Jianwen Chen, Xinyu Yang, Peng Xia, Arian Azarang, Yueh Z Lee, Gang Li, Hongtu Zhu, Yun Li, Beidi Chen, Huaxiu Yao

机构 * UNC-Chapel Hill(北卡罗来纳大学教堂山分校) Carnegie Mellon University(卡内基梅隆大学)

专题命中 临床大模型 :diagnosis(abstract);分类 cs.LG

AI总结 MedVerse通过将医学推理转化为可并行的有向无环图过程,提升复杂医疗问题的效率和可靠性,实验表明其在通用大语言模型上提升了8.9%的性能,同时减少推理延迟并提高生成吞吐量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14751 2026-04-17 cs.CL 50%

Query pipeline optimization for cancer patient question answering systems

癌症患者问答系统中的查询管道优化

Maolin He, Rena Gao, Mike Conway, Brian E. Chapman

机构 * School of Computing and Information Systems, University of Melbourne(墨尔本大学计算与信息系统学院) Health Data Science and Biostatistics, University of Texas Southwestern Medical Center(德克萨斯西南医学中心健康数据科学与生物统计学)

专题命中 临床大模型 :biomedical(abstract)

AI总结 本文提出了一种针对癌症患者问答系统的RAG查询管道三方面优化方法,通过改进文档检索、段落检索和语义表示,提升了回答准确性。

Comments This paper has been accepted as a Findings Paper in ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 诊断辅助 8 篇

2604.14259 2026-04-17 q-bio.TO cs.LG eess.IV 81%

Continual Learning for fMRI-Based Brain Disorder Diagnosis via Functional Connectivity Matrices Generative Replay

基于功能性连接矩阵生成性重放的fMRI脑部疾病诊断持续学习

Qianyu Chen, Shujian Yu

机构 * Nanyang Technological University(南洋理工大学) VU Amsterdam(阿姆斯特丹大学) UiT The Arctic University of Norway(北欧大学)

专题命中 诊断辅助 :diagnosis(title,abstract);分类 cs.LG、q-bio、eess.IV

AI总结 本文提出首个针对异构临床站点的fMRI诊断持续学习框架,通过生成对抗网络生成真实功能性连接矩阵,并结合多级知识蒸馏策略和分层上下文老虎机方案,有效缓解灾难性遗忘问题。

Comments manuscript accepted by CVPR 2026, code is available from \url{https://github.com/4me808/FORGE}

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09463 2026-04-17 cs.LG cs.AI 79%

Comorbidity-Informed Transfer Learning for Neuro-developmental Disorder Diagnosis

考虑共病的迁移学习用于神经发育障碍诊断

Xin Wen, Shijie Guo, Wenbo Ning, Rui Cao, Jie Xiang, Xiaobo Liu, Jintai Chen

机构 * School of Software, Taiyuan University of Technology(太原科技大学软件学院) School of Computer Science(Data Science),Taiyuan University of Technology(太原科技大学计算机科学(数据科学)学院)

专题命中 诊断辅助 :diagnosis(title,abstract);分类 cs.LG

AI总结 本文提出CITL框架,结合迁移学习和伪标签去除fMRI时间域干扰,生成新表示并用于分类,提升神经发育障碍诊断效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15066 2026-04-17 cs.LG cs.AI cs.MM 74%

Time-RA: Towards Time Series Reasoning for Anomaly Diagnosis with LLM Feedback

Time-RA:面向时间序列推理的异常诊断方法:基于LLM反馈

Yiyuan Yang, Zichuan Liu, Lei Song, Kai Ying, Zhiguang Wang, Tom Bamford, Svitlana Vyetrenko, Jiang Bian, Qingsong Wen

机构 * University of Oxford(牛津大学) Nanjing University(南京大学) MSRA(微软研究院) SJTU(上海交通大学) Abel AI Outsampler University of Strasbourg(斯特拉斯堡大学) Squirrel Ai Learning(Squirrel AI学习)

专题命中 诊断辅助 :diagnosis(title);分类 cs.LG

AI总结 本文提出Time-RA,通过引入RATs40K多模态数据集,改进时间序列异常检测的生成式推理方法,提升诊断准确性和解释性,实现可解释的多模态时间序列分析。

Comments ACL 2026 Findings. 27 pages, 11 figures, 15 tables. Code and dataset are publicly available

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.15190 2026-04-17 cs.AI cs.CL 71%

Meituan Merchant Business Diagnosis via Policy-Guided Dual-Process User Simulation

美团商户业务诊断 via 政策引导的双过程用户模拟

Ziyang Chen, Renbing Chen, Daowei Li, Jinzhi Liao, Jiashen Sun, Ke Zeng, Xiang Zhao

机构 * Independent Researcher(独立研究者)

专题命中 诊断辅助 :diagnosis(title)

AI总结 本文提出Policy-Guided Hybrid Simulation框架,通过双过程融合提升商户策略评估的准确性,实验显示其在美团101家商户中误差降低45.8%。

Comments 5 pages, 3 figures, 2 tables, accepted at SIGIR 2026 Industry Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14475 2026-04-17 cs.AI 71%

Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and Improve

Evo-MedAgent:超越单次诊断的具有记忆、反思和改进能力的代理

Weixiang Shen, Bailiang Jian, Jun Li, Che Liu, Johannes Moll, Xiaobin Hu, Daniel Rueckert, Hongwei Bran Li, Jiazhen Pan

机构 * Technical University of Munich(慕尼黑技术大学) TUM University Hospital(TUM大学医院) LMU Munich(慕尼黑大学) National University of Singapore(新加坡国立大学) Munich Center for Machine Learning(慕尼黑机器学习中心) Imperial College London(伦敦帝国理工学院)

专题命中 诊断辅助 :diagnosis(title)

AI总结 Evo-MedAgent通过引入自我进化记忆模块,使医疗代理在测试时实现跨案例学习,提升多选题准确率,无需额外训练,适用于任何冻结模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.15065 2026-04-17 cs.CV 57%

Learning Where to Embed: Noise-Aware Positional Embedding for Query Retrieval in Small-Object Detection

学习嵌入位置:面向小目标检测查询检索的噪声感知位置嵌入

Yangchen Zeng, Zhenyu Yu, Dongming Jiang, Wenbo Zhang, Yifan Hong, Zhanhua Hu, Jiao Luo, Kangning Cui

机构 * Southeast University(东南大学) Fudan University(复旦大学) The University of Texas at Dallas(德克萨斯大学达拉斯分校) Zhejiang Normal University(浙江师范大学) Data Space Research Institute, Hefei Comprehensive National Science Center(合肥综合国家科学中心数据空间研究院) Rice University(里士满大学) Huazhong Agricultural University(华中农业大学) City University of Hong Kong (Dongguan)(香港城市大学(东莞)) Wake Forest University(威克森林大学)

专题命中 诊断辅助 :diagnosis(abstract);分类 cs.CV

AI总结 本文提出HELP框架,通过选择性保留前景区域的位置编码来学习嵌入位置,结合热图引导的位置嵌入机制和线性蛇卷积提升小目标检索性能,实现参数减少59.4%的同时保持精度。

Comments Accepted to ACM ICMR 2026; 14 pages, 6 figures, and 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14209 2026-04-17 cs.LG cs.AI stat.ML 57%

Towards Verified and Targeted Explanations through Formal Methods

通过形式方法实现验证和定向的解释

Hanchen David Wang, Diego Manzanas Lopez, Preston K. Robinette, Ipek Oguz, Taylor T. Johnson, Meiyi Ma

机构 * Vanderbilt University(范德比大学)

专题命中 诊断辅助 :diagnosis(abstract);分类 cs.LG

AI总结 本文提出ViTaX框架,通过形式方法生成具有数学保证的定向半事实解释,解决传统XAI方法在可解释性和可信度上的不足。

Comments Paper has been accepted at JAIR

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22174 2026-04-17 cs.DC cs.AI cs.AR cs.CR cs.LG 57%

BitFlipScope: Scalable Fault Localization and Recovery for Bit-Flip Corruptions in LLMs

BitFlipScope:面向LLM中位翻故障的可扩展故障定位与恢复

Muhammad Zeeshan Karamat, Sadman Saif, Christiana Chamon Garcia

机构 * Bradly Dept. of ECE(布拉利电子工程系) Virginia Tech(弗吉尼亚理工学院)

专题命中 诊断辅助 :diagnosis(abstract);分类 cs.LG

AI总结 BitFlipScope通过两种部署场景实现LLM中位翻故障的定位与恢复,利用差分分析和残差路径扰动等方法,支持轻量级性能恢复,提升硬件环境下的模型可靠性。

Comments Accepted at the IEEE International Symposium on Hardware Oriented Security and Trust (HOST) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 病理影像 1 篇

2510.23838 2026-04-17 gr-qc astro-ph.CO hep-th 50%

Imperfect dark matter with higher derivatives

具有高阶导数的不完美暗物质

Mohammad Ali Gorji

专题命中 病理影像 :pathology(abstract)

AI总结 本文提出一种具有高阶导数的作用量,描述非理想流体的能量动量张量,通过高阶导数耦合可避免caustic奇点,解决仿真实体暗物质的路径问题。

Comments 17+6 pages, no figure, matches published version

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 医疗多模态 3 篇

2604.10410 2026-04-17 cs.AI 75%

CWCD: Category-Wise Contrastive Decoding for Structured Medical Report Generation

CWCD:用于结构化医学报告生成的类别级对比解码

Shantam Srivastava, Mahesh Bhosale, David Doermann, Mingchen Gao

机构 * The Department of Computer Science and Engineering(计算机科学与工程系) University at Buffalo, The State University of New York, NY, USA(布法罗大学,纽约州立大学)

专题命中 医疗多模态 :pathology(abstract);diagnosis(abstract);radiology(abstract)

AI总结 本文提出CWCD框架,通过类别特定参数化和对比正常与遮蔽X光片生成结构化放射报告,提升临床效果和生成质量。

Comments Accepted to MIDL 2026 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14866 2026-04-17 cs.CV cs.AI 57%

MetaDent: Labeling Clinical Images for Vision-Language Models in Dentistry

MetaDent: 临床图像标注用于牙科领域视觉-语言模型

Meng-Xun Li, Wen-Hui Deng, Zhi-Xing Wu, Chun-Xiao Jin, Jia-Min Wu, Yue Han, James Kit Hon Tsoi, Gui-Song Xia, Cui Huang

机构 * School of Computer Science, Wuhan University, Wuhan, Hubei, China(计算机科学学院,武汉大学,武汉,湖北,中国)

专题命中 医疗多模态 :medical image(abstract);分类 cs.CV

AI总结 本文提出MetaDent,通过大规模牙科图像数据集、半结构化标注框架和综合基准测试,解决牙科内窥镜图像分析中缺乏细粒度标注的问题,验证了VLMs在临床图像理解中的性能局限。

Comments Project website: https://menxli.github.io/metadent

Journal ref Journal of Dental Research, p.00220345261424242 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14656 2026-04-17 cs.AI cs.CL cs.CV 57%

Rethinking Patient Education as Multi-turn Multi-modal Interaction

重新思考患者教育作为多轮多模态交互

Zonghai Yao, Zhipeng Tang, Chengtao Lin, Xiong Luo, Benlu Wang, Juncheng Huang, Chin Siang Ong, Hong Yu

机构 * VA Bedford Health Care(VA贝德福德医疗中心) UMass Amherst(马萨诸塞大学阿默斯特分校) UMass Lowell(马萨诸塞大学洛厄尔分校) Yale University(耶鲁大学) National University of Singapore(新加坡国立大学) Yale School of Medicine(耶鲁医学院)

专题命中 医疗多模态 :radiology(abstract);分类 cs.CV

AI总结 本文提出MedImageEdu基准,通过多轮多模态交互提升患者教育效果,评估咨询过程和最终响应质量,发现多模态模型在视觉 grounding、安全性和情绪互动方面存在不足。

Comments Equal contribution for the first two authors

详情

展开后加载摘要…

URL PDF HTML 收藏