Comments9 pages, 5 figures. Empirical study of in-context learning and LoRA fine-tuning for synthetic tabular data generation, introducing the phenomenon of categorical prior lock-in. Under review
Unifying Adversarially Robust Model Experts in Vision-Language Models
统一视觉-语言模型中的对抗鲁棒模型专家
Nguyen Duc Thai, Junhao Dong, Sua Qi Rong, Hua Yu, Yew-Soon Ong
机构
*
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Center for Frontier AI Research, Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局前沿人工智能研究中心)
MBTI: A Multi-Branch Efficient Fine-Tuning Framework for Hyperspectral Image Classification with Foundation Models
MBTI:一种用于基于基础模型的高光谱图像分类的多分支高效微调框架
Mingzhen Xu, Haonan Guo, Di Wang, Yinghua Qu, Zhiliang Zhou, Lei Zhang, Huiwen Yao, Rui Zhao, Fengxiang Wang, Gang Wan, Bo Du, Liangpei Zhang
机构
*
School of Computer Science, Wuhan University(武汉大学计算机科学学院)
;
State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing, Wuhan University(武汉大学测绘遥感信息工程国家重点实验室)
;
China Petroleum Pipeline Engineering Corporation(中国石油管道局工程有限公司)
;
Hebei Key Laboratory of Underground Energy Storage Technology(河北省地下储能技术重点实验室)
;
No. 7 Oil Production Plant, Changqing Oilfield Branch, PetroChina Company Limited(中国石油天然气股份有限公司长庆油田分公司第七采油厂)
;
Faculty of Electrical Engineering and Computer Science, Ningbo University(宁波大学电气工程与计算机科学学院)
;
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机科学与技术学院)
;
School of Aerospace Information, Space Engineering University(航天工程大学航天信息学院)
;
Key Laboratory of Intelligent Processing and Applicati(智能处理与应用重点实验室)
TextGaze: Prompting Gaze Target Estimation with Textual Scene Cues
TextGaze:利用文本场景线索提示注视目标估计
Junhui She, Fei Wang, Kun Li, Yiqi Nie, Yuxin Liu, Zhangling Duan, Xun Yang
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Hefei University of Technology(合肥工业大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
;
Anhui University(安徽大学)
;
United Arab Emirates University(阿联酋大学)
机构
*
School of Information Science and Technology, Northwest University(西北大学信息科学与技术学院)
;
Laboratory of Intelligent Information Processing, Institute of Computing Technology(中国科学院计算技术研究所智能信息处理重点实验室)
VLM-CASE: Vision-Language Model Enabled Context-Adaptive Safety Envelopes for Anticipatory Safe Autonomous Driving
VLM-CASE:面向前瞻安全自动驾驶的视觉语言模型赋能上下文自适应安全包络
Tianjia Yang, Ke Li, Ruwen Qin, Xianbiao Hu
机构
*
Department of Civil and Environmental Engineering, The Pennsylvania State University(宾夕法尼亚州立大学土木与环境工程系)
;
Department of Civil Engineering, Stony Brook University(石溪大学土木工程系)
Evaluating Vision-Language Models as a Zero-Shot Learning Alternative to You Only Look Once and Optical Character Recognition for Nigerian License Plate Recognition
评估视觉语言模型作为尼日利亚车牌识别的零样本学习替代方案:对比YOLO与光学字符识别
Ismail Ismail Tijjani, Ahmad Abubakar Mustapaha, Sunusi Ibrahim Muhammad, Muhammad Bashir Aliyu