Comments14 pages, 1 figure, 7 tables. Accepted to the 15th Workshop on Computational Approaches to Subjectivity, Sentiment & Social Media Analysis (WASSA) at EACL 2026, Rabat, Morocco
Journal refProceedings of the 15th Workshop on Computational Approaches to Subjectivity, Sentiment & Social Media Analysis (WASSA), 2026
机构
*
National Yang Ming Chiao Tung University(国立阳明交通大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
AI Research Center, Hon Hai Research Institute(鸿海研究院人工智能研究中心)
机构
*
School of Advanced Technology, Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学先进科技学院)
;
Department of Computer Science, University of Liverpool(利物浦大学计算机科学系)
;
Information Hub, Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)信息中心)
;
University of Southern California(南加州大学)
;
Duke Kunshan University(杜克大学昆山分校)
Widget2Code: From Visual Widgets to UI Code via Multimodal LLMs
Widget2Code: 通过多模态大语言模型从视觉小部件到UI代码
Houston H. Zhang, Tao Zhang, Baoze Lin, Yuanqi Xue, Yincheng Zhu, Huan Liu, Li Gu, Linfeng Ye, Ziqiang Wang, Xinxin Zuo, Yang Wang, Yuanhao Yu, Zhixiang Chi
机构
*
McMaster University(麦克马斯特大学)
;
University of Toronto(多伦多大学)
;
University of Waterloo(滑铁卢大学)
;
Concordia University(康考迪亚大学)
专题命中
GUI与屏幕智能体
:multimodal large language model(abstract);分类 cs.CV
Abhay Deshpande, Maya Guru, Rose Hendrix, Snehal Jauhri, Ainaz Eftekhar, Rohun Tripathi, Max Argus, Jordi Salvador, Haoquan Fang, Matthew Wallingford, Wilbert Pumacay, Yejin Kim, Quinn Pfeifer, Ying-Chun Lee, Piper Wolters, Omar Rayyan, Mingtong Zhang, Jiafei Duan, Karen Farley, Winson Han, Eli Vanderbilt, Dieter Fox, Ali Farhadi, Georgia Chalvatzaki, Dhruv Shah, Ranjay Krishna
机构
*
Allen Institute for AI(艾伦人工智能研究所)
;
University of Washington(华盛顿大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Princeton University(普林斯顿大学)
机构
*
School of Computer Science and Artificial Intelligence, Wuhan University of Technology(武汉理工大学计算机科学与人工智能学院)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
专题命中
幻觉与鲁棒性
:multimodal large language model(title,abstract);分类 cs.CV
机构
*
Indian Institute of Technology Jodhpur(印度理工学院焦特布尔分校)
;
AIM Intelligence
;
National Institute of Technology Agartala(国立阿加尔塔拉理工学院)
;
Fordham University(福特汉姆大学)
;
University of Fukui(福井大学)
;
Independent Researcher(独立研究员)
VOLMO: Versatile and Open Large Models for Ophthalmology
VOLMO:面向眼科学的多功能和开放型大模型
Zhenyue Qin, Younjoon Chung, Elijah Lee, Wanyue Feng, Xuguang Ai, Serina Applebaum, Minjie Zou, Yang Liu, Pan Xiao, Mac Singer, Amisha Dave, Aidan Gilson, Tiarnan D. L. Keenan, Emily Y. Chew, Zhiyong Lu, Yih-Chung Tham, Ron Adelman, Luciano V. Del Priore, Qingyu Chen
机构
*
Department of Biomedical Informatics & Data Science, Yale University(耶鲁大学生物医学信息学与数据科学系)
;
Ray and Stephanie Lane Computational Biology Department, Carnegie Mellon University(卡内基梅隆大学雷和斯蒂芬妮·兰德计算生物学系)
;
Yong Loo Lin School of Medicine, National University of Singapore(新加坡国立大学杨洛林医学院)
;
Department of Radiology, Washington University in Saint Louis(圣路易斯华盛顿大学放射科)
;
National Eye Institute, National Institutes of Health(国家卫生研究院眼科研究所)
;
National Library of Medicine, National Institutes of Health(国家卫生研究院国家医学图书馆)
专题命中
VLM训练与架构
:LLaVA(abstract);InternVL(abstract);multimodal large language model(abstract);MLLM(abstract)
Huy Hoang Nguyen, Cédric Jung, Shirin Salehi, Tobias Glück, Anke Schmeink, Andreas Kugi
机构
*
AIT Austrian Institute of Technology(奥地利理工学院)
;
Automation & Control Institute, Technical University of Vienna(维也纳技术大学自动化与控制研究所)
;
Chair of Information Theory and Data Analytics (INDA), RWTH Aachen University(亚琛工业大学信息理论与数据分析教席)
GIFT: Global Irreplaceability Frame Targeting for Efficient Video Understanding
GIFT:全局不可替代性帧目标用于高效视频理解
Junpeng Ma, Sashuai Zhou, Guanghao Li, Xin Gao, Yue Cao, Hengyu Zeng, Yuxiang Yan, Zhibin Wang, Jun Song, Bo Zheng, Shanghang Zhang, Jian Pu
机构
*
Institute of Science and Technology for Brain-inspired Intelligence, Fudan University(复旦大学类脑智能科学与技术研究院)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(北京大学计算机学院多媒体信息处理国家重点实验室)
;
Zhejiang University(浙江大学)
;
Alibaba Group Holding Limited(阿里巴巴集团控股有限公司)
;
Future Living Lab of Alibaba(阿里巴巴未来生活实验室)
CommentsSignificantly extended version of earlier work, with additional experiments, expanded discussion and related work, new experiments with human participants, and broader evaluation across multiple vision-language models