CommentsThe results presented in this paper are preliminary. Please note that the experiments are currently ongoing, and the final data is subject to change upon the completion of the study. All ideas, results, methods, and any content herein are the sole property of the authors
ImpedanceDiffusion: Diffusion-Based Global Path Planning for UAV Swarm Navigation with Generative Impedance Control
阻抗扩散:基于扩散的全局路径规划用于无人机群导航的生成阻抗控制
Faryal Batool, Yasheerah Yaqoot, Muhammad Ahsan Mustafa, Roohan Ahmed Khan, Aleksey Fedoseev, Dzmitry Tsetserukou
机构
*
Intelligent Space Robotics Laboratory, Center for Digital Engineering, Skolkovo Institute of Science and Technology(智能空间机器人实验室,数字工程中心,斯克尔科夫科学与技术研究所)
LLaVAShield: Safeguarding Multimodal Multi-Turn Dialogues in Vision-Language Models
LLaVAShield: 保障视觉语言模型中的多模态多轮对话安全
Guolei Huang, Qinzhi Peng, Gan Xu, Yao Huang, Yuxuan Lu, Yongjun Shen
机构
*
Southeast University(东南大学)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
;
Zhejiang University of Technology(浙江工业大学)
;
Tsinghua University(清华大学)
;
RealAI
Mitigating Long-Tail Bias in HOI Detection via Adaptive Diversity Cache
通过自适应多样性缓存缓解HOI检测中的长尾偏差
Yuqiu Jiang, Xiaozhen Qiao, Yifan Chen, Ye Zheng, Zhe Sun, Xuelong Li
机构
*
College of Future Information Technology, Fudan University(复旦大学未来信息技术学院)
;
School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)
;
Institute of Artificial Intelligence (TeleAI), China Telecom(中国电信人工智能研究院(TeleAI))
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
Fudan University(复旦大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
South China University of Technology(华南理工大学)
;
Nanjing University(南京大学)
;
Xiamen University(厦门大学)
;
CUHK MMLab(香港中文大学MMLab)
专题命中
VLM训练与架构
:InternVL(title,abstract);multimodal large language model(abstract);MLLM(abstract);分类 cs.CV
Fairness-Aware Fine-Tuning of Vision-Language Models for Medical Glaucoma Diagnosis
面向医疗青光眼诊断的公平性感知视觉语言模型微调
Zijian Gu, Yuxi Liu, Zhenhao Zhang, Song Wang
机构
*
Department of Computer Science, University of Rochester, NY, USA(罗切斯特大学计算机科学系)
;
Biostatistics and Health Data Science, School of Medicine, Indiana University, Indianapolis, IN, USA(印第安纳大学医学院生物统计学与健康数据科学系)
;
Department of Computer Science, University of Central Florida, FL, USA(佛罗里达州立大学计算机科学系)
VLCE: A Knowledge-Enhanced Framework for Image Description in Disaster Assessment
VLCE:一种用于灾害评估图像描述的知识增强框架
Md. Mahfuzur Rahman, Kishor Datta Gupta, Marufa Kamal, Fahad Rahman, Sunzida Siddique, Ahmed Rafi Hasan, Mohd Ariful Haque, Roy George
机构
*
Clark Atlanta University(克拉克亚特兰大大学)
;
BRAC University(布拉克大学)
;
United International University(国际联合大学)
;
Daffodil International University(花王国际大学)
V-Attack: Targeting Disentangled Value Features for Controllable Adversarial Attacks on LVLMs
V-Attack:针对解耦价值特征的可控对抗攻击LVLMs
Sen Nie, Jie Zhang, Jianxin Yan, Shiguang Shan, Xilin Chen
机构
*
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(中国科学院人工智能安全国家重点实验室,计算技术研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Zhejiang University(浙江大学)