Jolia: Concept-Level Vision-Language Alignment for 3D CT Contrastive Learning
Jolia: 用于3D CT对比学习的概念级视觉-语言对齐
Julien Khlaut, Charles Corbière, Baptiste Callard, Amaury Prat, Leo Butsanets, Antoine Saporta, Théo Danielou, Leo Machado, Korentin Le Floch, Tom Boeken, Pierre Manceron, Corentin Dancette
机构
*
Raidium
;
Department of Vascular and Oncological Interventional Radiology, Hôpital Européen Georges Pompidou, AP-HP(欧洲乔治·蓬皮杜医院血管与肿瘤介入放射科,AP-HP)
;
Faculté de Santé, Université Paris-Cité(巴黎西岱大学健康学院)
;
HEKA, INRIA(HEKA,法国国家信息与自动化研究所)
;
Imaging Department, Fondation Ophtalmologique Adolphe de Rothschild(阿道夫·罗斯柴尔德眼科基金会影像科)
Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling
面向多模态大语言模型的长视频快速有效理解:自适应准高斯采样
Kun Zhang, Chenxin Fang, Tao Chen, Baiyang Song, Yunhang Shen, Yiyi Zhou, Rongrong Ji
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(厦门大学多媒体可信感知与高效计算教育部重点实验室)
SurgAtlas: A Large-Scale Surgical Video-Language Dataset with 2,391 Hours of Open and Minimally Invasive Surgery
SurgAtlas:一个包含2,391小时开放和微创手术的大规模手术视频-语言数据集
Filippos Bellos, Andre S. Gala-Garza, Miaowei Wang, Alyssa M. Hardin, Ahmad M. Hider, Yayuan Li, Jing Bi, Susan Liang, Chenliang Xu, Donald S. Likosky, Jason J. Corso
机构
*
University of Michigan(密歇根大学)
;
University of Edinburgh(爱丁堡大学)
;
Vanderbilt University Medical Center(范德比尔特大学医学中心)
;
University of Colorado(科罗拉多大学)
;
University of Rochester(罗切斯特大学)
;
Michigan Medicine(密歇根医学)
Spatio-Temporal Mixture-of-Modality-Experts Diffusion for Quantitative DCE-MRI Synthesis from Incomplete MR Sequences
时空模态专家混合扩散:从不完整MR序列合成定量DCE-MRI
Junhyeok Lee, Kyu Sung Choi
机构
*
Interdisciplinary Program in Cancer Biology, Seoul National University College of Medicine(首尔大学医学院癌症生物学跨学科项目)
;
Department of Radiology, Seoul National University Hospital(首尔大学医院放射科)
;
Department of Radiology, Seoul National University College of Medicine(首尔大学医学院放射科)
;
Healthcare AI Research Institute, Seoul National University Hospital(首尔大学医院医疗人工智能研究所)
Qinzhe Yang, Chenyang Liu, Jia Xu, Zhenwei Shi, Zhengxia Zou
机构
*
Shen Yuan Honors College, Beihang University(北航沈元荣誉学院)
;
Department of Aerospace Intelligent Science and Technology, School of Astronautics, Beihang University(北京航空航天大学宇航学院航天智能科学与技术系)
;
State Key Laboratory of Virtual Reality Technology and Systems, Beihang University(北京航空航天大学虚拟现实技术与系统国家重点实验室)
;
Qian Xuesen Laboratory of Space Technology, China Academy of Space Technology(中国空间技术研究院钱学森空间技术实验室)
When Multi-Sensor Fusion Fails to Generalize: Cattle Posture Classification Under Animal-Level and Temporal Distribution Shift
当多传感器融合无法泛化:动物级别和时间分布偏移下的牛姿势分类
Leutrim Uka, Severino Pinto, Gundula Hoffmann, Marina M. -C. Höhne
机构
*
Institute of Computer Science, University of Potsdam(波茨坦大学计算机科学研究所)
;
Department of Sensors and Modelling, Leibniz Institute for Agricultural Engineering and Bioeconomy - ATB(莱布尼茨农业工程与生物经济研究所传感器与建模系)
;
Department of Data Science in Bioeconomy, Leibniz Institute for Agricultural Engineering and Bioeconomy - ATB(莱布尼茨农业工程与生物经济研究所生物经济数据科学系)
An iterative energy-based multimodal transformer for joint retrieval of wheat soil moisture, leaf area index, and plant height from Sentinel-1 and Sentinel-2 time series
Shubham Kumar Singh, Peilei Fan, Suraj A. Yadav, Rajendra Prasad, Prashant K Srivastava
机构
*
Department of Urban and Environmental Policy and Program, Tufts University(Tufts大学城市与环境政策与项目系)
;
Electrical and Computer Engineering Department, Mississippi State University(密苏里州立大学电气与计算机工程系)
;
Department of Physics, Indian Institute of Technology (BHU)(印度理工学院(BHU)物理系)
;
Institute of Environment and Sustainable Development, Banaras Hindu University(巴纳尔斯赫尔大学环境与可持续发展研究所)
Point Cloud Diffusion with Global and Local Reconstruction for Instance-Level 3D Anomaly Detection
基于全局与局部重建的点云扩散用于实例级3D异常检测
Linchun Wu, Qin Zou, Jiwen Lu, Qingquan Li
机构
*
School of Computer Science, Wuhan University(武汉大学计算机学院)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
Guangdong Artificial Intelligence and Digital Economy Laboratory (SZ)(广东人工智能与数字经济实验室(深圳))
VPA-Guard: Defending and Benchmarking Image-to-Video Generation Against Visual Prompt Attacks
VPA-Guard:防御与基准测试图像到视频生成中的视觉提示攻击
Yining Sun, Haoyu Kang, Jiajun Wu, Heng Zhang, Danyang Zhang, Zhenjun Zhao, Haochen Han, Fangming Liu, Wai Kin Victor Chan, Alex Jinpeng Wang
机构
*
Tsinghua University(清华大学)
;
Central South University(中南大学)
;
South China Normal University(华南师范大学)
;
ByteDance Inc.(字节跳动公司)
;
Pengcheng Laboratory(鹏城实验室)
Stage-Aware and Roughness-Constrained Diffusion Policy for Multi-Stage Robotic Polishing
阶段感知与粗糙度约束的扩散策略用于多阶段机器人抛光
Shuai Ke, Jiexin Zhang, Huan Zhao, Zhiao Wei, Yikun Guo, Tiange Wu, Guoqiang Guo, Haoyuan Zhou, Jie Pan, Han Ding
机构
*
State Key Laboratory of Intelligent Manufacturing Equipment and Technology, Huazhong University of Science and Technology(华中科技大学智能制造装备与技术国家重点实验室)
;
Shanghai Spaceflight Precision Machinery Institute(上海航天精密机械研究所)