Quality Text, Robust Vision: The Role of Language in Enhancing Visual Robustness of Vision-Language Models
高质量文本,稳健视觉:语言在增强视觉语言模型视觉稳健性中的作用
Futa Waseda, Saku Sugawara, Isao Echizen
机构
*
The University of Tokyo(东京大学)
;
National Institute of Informatics(日本信息处理研究所)
;
The University of Tokyo, National Institute of Informatics(东京大学、日本信息处理研究所)
Joint Lossless Compression and Steganography for Medical Images via Large Language Models
通过大语言模型实现医学图像的联合无损压缩与隐写术
Pengcheng Zheng, Xiaorong Pu, Kecheng Chen, Jiaxin Huang, Meng Yang, Bai Feng, Yazhou Ren, Jianan Jiang, Chaoning Zhang, Yang Yang, Heng Tao Shen
机构
*
Center for Future Media and School of Computer Science and Engineering, University of Electronic Science and Technology of China(未来媒体中心和电子科技大学计算机科学与工程学院)
;
Department of Computer Science and Engineering, University of Electronic Science and Technology of China(计算机科学与工程学院,电子科技大学)
;
Department of Electrical Engineering, and the Center for Intelligent Multidimensional Data Analysis, City University of Hong Kong(电子工程系和智能多维数据分析中心,城市大学)
;
Department of Machine Learning, Mohamed bin Zayed University of Artificial Intelligence(机器学习系,Mohamed bin Zayed人工智能大学)
专题命中
指令微调
:large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
Layer-wise LoRA fine-tuning: a similarity metric approach
逐层LoRA微调:相似性度量方法
Keith Ando Ogawa, Bruno Lopes Yamamoto, Lucas Lauton de Alcantara, Lucas Pellicer, Rosimeire Pereira Costa, Edson Bollis, Anna Helena Reali Costa, Artur Jordao
专题命中
指令微调
:large language model(abstract);language model(abstract);分类 cs.LG
Self-Guided Process Reward Optimization with Redefined Step-wise Advantage for Process Reinforcement Learning
基于重新定义的逐步优势的自引导过程奖励优化用于过程强化学习
Wu Fei, Shuxian Liang, Yibo Yang, Yang Lin, Jing Tang, Lei Chen, Xiansheng Hua, Hao Kong
机构
*
Terminus Group(Terminus集团)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
King Abdullah University of Science and Technology(国王阿卜杜勒阿齐兹大学)
专题命中
后训练与偏好优化
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
机构
*
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室)
;
School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学交叉科学学院)
;
Tsinghua University(清华大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation
PixDLM:一种用于无人机推理分割的双路径多模态语言模型
Shuyan Ke, Yifan Mei, Changli Wu, Yonghan Zheng, Jiayi Ji, Liujuan Cao, Rongrong Ji
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
当伦理与收益相悖时:道德两难情境下的大语言模型智能体
Steffen Backmann, David Guzman Piedrahita, Terry Jingchen Zhang, Emanuel Tewolde, Rada Mihalcea, Bernhard Schölkopf, Zhijing Jin
机构
*
ETH Zürich(苏黎世联邦理工学院)
;
University of Zurich(苏黎世大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of Michigan(密歇根大学)
;
Max Planck Institute for Intelligent Systems, Tübingen(图宾根人工智能研究所)
;
University of Toronto(多伦多大学)
;
Vector Institute(向量研究所)
机构
*
Department of Electrical and Computer Engineering, Virginia Tech(弗吉尼亚理工学院电气与计算机工程系)
;
Department of Computer Science, Brown University(布朗大学计算机科学系)
;
Department of Computer Science, University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校计算机科学系)
;
Khoury College of Computer Sciences, Northeastern University(东北大学科里尔计算机科学学院)
专题命中
推理与问题求解
:large language model(abstract);language model(abstract);分类 cs.AI
CommentsThis version of the contribution has been accepted for publication at ICAI 2026, after peer review but is not the Version of Record and does not reflect post-acceptance improvements, or any corrections
机构
*
University of British Columbia(英属哥伦比亚大学)
;
Johns Hopkins University(约翰·霍普金斯大学)
;
National University of Singapore(新加坡国立大学)
;
Columbia University(哥伦比亚大学)
;
University of California, Los Angeles(加利福尼亚大学洛杉矶分校)
;
Style3D(无合适对应中文名)