TextME: Bridging Unseen Modalities Through Text Descriptions
TextME: 通过文本描述弥合未见模态
Soyeon Hong, Jinchan Kim, Jaegook You, Seungtaek Choi, Suha Kwak, Hyunsouk Cho
机构
*
Department of Artificial Intelligence, Ajou University, Suwon, South Korea
;
Division of Language \& AI, Hankuk University of Foreign Studies, Seoul, Korea
;
Graduate School of AI, POSTECH, Pohang, Korea
;
Department of Software, Ajou University, Suwon, South Korea
机构
*
School of Computer Science, University of Nottingham Ningbo(计算机科学学院,诺丁汉大学宁波分校)
;
School of Artificial Intelligence, Shenzhen University(人工智能学院,深圳大学)
MultiDiffSense: Diffusion-Based Multi-Modal Visuo-Tactile Image Generation Conditioned on Object Shape and Contact Pose
MultiDiffSense: 基于扩散的多模态视觉-触觉图像生成,基于物体形状和接触姿态
Sirine Bhouri, Lan Wei, Jian-Qing Zheng, Dandan Zhang
机构
*
Department of Bioengineering, Imperial-X Initiative, Imperial College London(生物工程系、Imperial-X计划、帝国理工学院伦敦分校)
;
CAMS-Oxford Institute, University of Oxford(CAMS-牛津研究所、牛津大学)
AuditoryHuM: Auditory Scene Label Generation and Clustering using Human-MLLM Collaboration
AuditoryHuM: 利用人机协同生成和聚类听觉场景标签
Henry Zhong, Jörg M. Buchholz, Julian Maclaren, Simon Carlile, Richard F. Lyon
机构
*
Australian Hearing Hub, Macquarie University, Sydney, Australia(澳大利亚听力中心、麦觉里大学、悉尼、澳大利亚)
;
Google Research Australia, Sydney, Australia(谷歌澳大利亚研究、悉尼、澳大利亚)
Efficient endometrial carcinoma screening via cross-modal synthesis and gradient distillation
通过跨模态合成与梯度蒸馏实现高效的子宫内膜癌筛查
Dongjing Shan, Yamei Luo, Jiqing Xuan, Lu Huang, Jin Li, Mengchu Yang, Zeyu Chen, Fajin Lv, Yong Tang, Chunxiang Zhang
机构
*
School of Medical Information and Engineering, Southwest Medical University(西南医科大学医学信息与工程学院)
;
School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院)
;
Department of Ultrasound, Affiliated Hospital of Southwest Medical University(西南医科大学附属医院超声科)
;
Department of Functional Examination Unit, Zibo Hospital of Traditional Chinese Medicine(淄博中医药医院功能检查科)
;
Key Laboratory of Medical Electrophysiology, Ministry of Education& Medical Electrophysiological Key Laboratory of Sichuan Province, Institute of Cardiovascular Research, Southwest Medical University(教育部医学电生理重点实验室、四川省医学电生理重点实验室、西南医科大学心血管研究所)
;
Department of Radiology, the First Affiliated Hospital of Chongqing Medical University(重庆医科大学第一附属医院放射科)
;
Department of Cardiology, Affiliated Hospital of Southwest Medical University(西南医科大学附属医院心内科)
;
International Research Center for Complexity Sciences, Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院复杂科学研究中心)
;
Institute of Intelligent Chinese Medicine, Chongqing University of Chinese Medicine(重庆中医药大学智能中药研究院)
;
Basic Medicine Research Innovation Center for Cardiometabolic Diseases, Ministry of Education, Southwest Medical University(教育部心脑血管疾病基础医学研究创新中心、西南医科大学)
JavisDiT++: Unified Modeling and Optimization for Joint Audio-Video Generation
JavisDiT++:联合音频视频生成的统一建模与优化
Kai Liu, Yanhao Zheng, Kai Wang, Shengqiong Wu, Rongjunchen Zhang, Jiebo Luo, Dimitrios Hatzinakos, Ziwei Liu, Hao Fei, Tat-Seng Chua
机构
*
Zhejiang University(浙江大学)
;
National University of Singapore(新加坡国立大学)
;
University of Toronto(多伦多大学)
;
HiThink Research(HiThink研究)
;
University of Rochester(罗切斯特大学)
;
Nanyang Technological University(南洋理工大学)
机构
*
Graduate School of Information Science and Engineering, Ritsumeikan University, Japan(立命馆大学信息科学与工程研究生院)
;
Japan Advanced Institute of Science and Technology (JAIST)(日本先进科学研究院)
;
College of Information Science and Engineering, Ritsumeikan University, Japan(立命馆大学信息科学与工程学院)
Setting up for failure: automatic discovery of the neural mechanisms of cognitive errors
失败的设定:自动发现认知错误的神经机制
Puria Radmard, Paul M. Bays, Máté Lengyel
机构
*
Department of Engineering, University of Cambridge(工程系,剑桥大学)
;
Department of Psychology, University of Cambridge(心理学系,剑桥大学)
;
Department of Cognitive Science, Central European University(认知科学系,中央欧亚大学)
Minfeng Zhu, Zi Wang, Sizhe Ji, Zhengtong Du, Shengqiang Tai, Junming Ke, Xiao Deng, Zanlang Yin, Xiuqi Huang, Heyu Wang, Wei Chen
机构
*
State Key Lab of CAD&CG, Zhejiang University(浙江大学计算机辅助设计与图形学国家重点实验室)
;
Polytechnic Institute, Zhejiang University(浙江大学多科大学院)
;
Hangzhou Research Institute of AI and Holographic Technology(杭州人工智能与全息技术研究 institutes)
;
Volkswagen Group Innovation(大众集团创新)
;
School of Mathematical Science, Zhejiang University(浙江大学数学科学学院)
HIME: Mitigating Object Hallucinations in LVLMs via Hallucination Insensitivity Model Editing
HIME: 通过幻觉不敏感模型编辑缓解LVLMs中的物体幻觉
Ahmed Akl, Abdelwahed Khamis, Ali Cheraghian, Zhe Wang, Sara Khalifa, Kewen Wang
机构
*
School of Information and Communication Technology, Griffith University, Australia(信息与通信技术学院,格里菲斯大学)
;
Data61, CSIRO, Australia(Data61,澳大利亚联邦科学与工业研究组织)
;
School of Engineering, Macquarie University, Sydney, Australia(工程学院,麦觉大学)
;
School of Information Systems, Queensland University of Technology, Australia(信息系统学院,昆士兰技术大学)
Developing a Multi-Agent System to Generate Next Generation Science Assessments with Evidence-Centered Design
开发一个多智能体系统以生成下一代科学评估并采用证据中心设计
Yaxuan Yang, Jongchan Park, Yifan Zhou, Xiaoming Zhai
机构
*
AI4STEM Education Center, University of Georgia(AI4STEM教育中心,佐治亚大学)
;
Department of Educational Psychology, University of Georgia(教育心理学系,佐治亚大学)
;
School of Computing, University of Georgia(计算学院,佐治亚大学)
;
Department of Mathematics, Science, and Social Studies Education, University of Georgia(数学、科学与社会科学教育系,佐治亚大学)
机构
*
Peking University(北京大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Nankai University(南开大学)
;
Beijing Institute of Technology(北京理工大学)
;
Baichuan Inc.(百度文心)
机构
*
Shenzhen Technology University(深圳技术大学)
;
Shenzhen University of Information Technology(深圳信息科技学院)
;
University of Washington(华盛顿大学)
;
Foundation Model Team, Meituan(美团基础模型团队)
;
National University of Singapore(新加坡国立大学)
;
Tongji University(同济大学)
FOCA: Frequency-Oriented Cross-Domain Forgery Detection, Localization and Explanation via Multi-Modal Large Language Model
FOCA:基于多模态大语言模型的频率导向跨域伪造检测、定位与解释
Zhou Liu, Tonghua Su, Hongshi Zhang, Fuxiang Yang, Donglin Di, Yang Song, Lei Fan
机构
*
Harbin Institute of Technology(哈尔滨工业大学)
;
DZ-Matrix
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy(广东省人工智能与数字经济实验室)
;
Chongqing Research Institute of HIT(哈尔滨工业大学重庆研究院)
;
University of New South Wales(新南威尔士大学)
Towards Personalized Multi-Modal MRI Synthesis across Heterogeneous Datasets
面向异质数据集的个性化多模态MRI合成
Yue Zhang, Zhizheng Zhuo, Siyao Xu, Shan Lv, Zhaoxi Liu, Jun Qiu, Qiuli Wang, Yaou Liu, S. Kevin Zhou
机构
*
University of Science and Technology of China(中国科学技术大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Capital Medical University(首都医科大学)
;
Army Medical University(中国人民解放军陆军军医大学)
;
China National Clinical Research Center for Neurological Diseases(中国国家神经疾病临床研究中心)
BriMA: Bridged Modality Adaptation for Multi-Modal Continual Action Quality Assessment
BriMA:基于多模态适应的持续动作质量评估
Kanglei Zhou, Chang Li, Qingyi Pan, Liyuan Wang
机构
*
Department of Psychological and Cognitive Sciences, 2 Department of Statistics and Data Science, Tsinghua University(1 心理学与认知科学系,2 统计学与数据科学系,清华大学)