Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise
基于符号的优化器在重尾噪声下表现有效
Dingzhi Yu, Hongyi Tao, Yuanyu Wan, Luo Luo, Lijun Zhang
机构
*
State Key Laboratory of Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
;
School of Software Technology, Zhejiang University(浙江大学软件学院)
;
School of Data Science, Fudan University(复旦大学数据科学学院)
机构
*
Shanghai Key Laboratory of Data Science, College of Computer Science and Artificial Intelligence, Fudan University(上海数据科学 key laboratory,计算机科学与人工智能学院,复旦大学)
;
School of Data Science, Fudan University(数据科学学院,复旦大学)
;
Ant Group(蚂蚁集团)
See Tomorrow, Act Today: Foresight-Driven Autonomous Driving
预见未来,立即行动:基于预见的自动驾驶
Bozhou Zhang, Nan Song, Yuang Wang, Jiankang Deng, Xiatian Zhu, Li Zhang
机构
*
School of Data Science, Fudan University(复旦大学数据科学学院)
;
Shanghai Innovation Institute(上海创新研究院)
;
Imperial College London(伦敦帝国理工学院)
;
University of Surrey(萨里大学)
Topology-Enhanced Alignment for Large Language Models: Trajectory Topology Loss and Topological Preference Optimization
基于拓扑的大型语言模型对齐:轨迹拓扑损失与拓扑偏好优化
Yurui Pan, Ke Xu, Bo Peng
机构
*
School of Computing and Intelligent Innovation, Fudan University(复旦大学计算与智能创新学院)
;
School of Economics and Management, Tongji University(同济大学经济与管理学院)
;
College of Information Technology, Shanghai Ocean University(上海海洋大学信息学院)
Towards Security-Auditable LLM Agents: A Unified Graph Representation
迈向安全可审计的LLM代理:一种统一的图表示
Chaofan Li, Lyuye Zhang, Jintao Zhai, Siyue Feng, Xichun Yang, Huahao Wang, Shihan Dou, Yu Ji, Yutao Hu, Yueming Wu, Yang Liu, Deqing Zou
机构
*
Huazhong University of Science and Technology(华中科技大学)
;
Nanyang Technological University(南洋理工大学)
;
Fudan University(复旦大学)
;
Chongqing University of Posts and Telecommunications(重庆邮电大学)
;
Donghua University(东华大学)
Data Augmentation of Contrastive Learning is Estimating Positive-incentive Noise
对比学习中的数据增强是估计正激励噪声
Hongyuan Zhang, Yanchen Xu, Sida Huang, Xuelong Li
机构
*
The University of Hong Kong(香港大学)
;
Institute of Artificial Intelligence (TeleAI), China Telecom(中国电信人工智能研究院)
;
Fudan University(复旦大学)
;
School of Artificial Intelligence, OPtics(人工智能学院)
;
ElectroNics (iOPEN), Northwestern Polytechnical University(西北工业大学电子学院)
机构
*
Fudan University(复旦大学)
;
Tongji University(同济大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
The University of Hong Kong(香港大学)
;
University of California San Diego(加州大学圣地亚哥分校)
;
Nanjing University of Posts and Telecommunications(南京邮电大学)
ZScribbleSeg: A comprehensive segmentation framework with modeling of efficient annotation and maximization of scribble supervision
ZScribbleSeg: 一种综合分割框架,包含高效标注建模和scribble监督最大化
Ke Zhang, Bomin Wang, Hangqi Zhou, Xiahai Zhuang
机构
*
School of Data Science, Fudan University, Shanghai, 200433, China(复旦大学数据科学学院)
;
Department of Electrical and Computer Engineering, Johns Hopkins University, Baltimore, USA(约翰霍普金斯大学电气与计算机工程系)
EA-WM: Event-Aware Generative World Model with Structured Kinematic-to-Visual Action Fields
EA-WM:事件感知生成世界模型与结构运动-视觉动作场
Zhaoyang Yang, Yurun Jin, Lizhe Qi, Cong Huang, Kai Chen
机构
*
Fudan University(复旦大学)
;
Zhongguancun Academy(中关村学院)
;
Zhongguancun Institute of Artificial Intelligence(中关村人工智能研究院)
;
University of Science and Technology of China(中国科学技术大学)
;
DeepCybo(深瞳)
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
King’s College London(伦敦国王学院)
;
East China University of Science and Technology(东华大学)
;
East China Normal University(东华师范大学)
;
Tongji University(同济大学)
;
Peking University(北京大学)
;
Fudan University(复旦大学)
;
University of Science and Technology of China(中国科学技术大学)
Hidden in the Multiplicative Interaction: Uncovering Fragility in Multimodal Contrastive Learning
乘积交互中的隐藏问题:揭示多模态对比学习中的脆弱性
Tillmann Rheude, Stefan Hegselmann, Roland Eils, Benjamin Wild
机构
*
Berlin Institute of Health, Charité - Universitätsmedizin Berlin(柏林健康研究所,柏林查理医院)
;
Intelligent Medicine Institute, Fudan University(智能医学研究院,复旦大学)
;
Department of Mathematics and Computer Science, Freie Universität Berlin(数学与计算机科学系,柏林自由大学)
FRISM: Fine-Grained Reasoning Injection via Subspace-Level Model Merging for Vision-Language Models
FRISM:通过子空间级模型融合实现细粒度推理注入用于视觉-语言模型
Chenyu Huang, Peng Ye, Xudong Tan, Jinhan Mu, Shenghe Zheng, Li Shen, Tao Chen
机构
*
College of Future Information Technology, Fudan University, Shanghai, China(复旦大学未来信息科技学院,中国)
;
Shanghai Innovation Institute, China(上海创新研究院,中国)
;
The Chinese University of Hong Kong, China(香港中文大学,中国)
;
Harbin Institute of Technology, China(哈尔滨工业大学,中国)
;
Shanghai Artificial Intelligence Laboratory, China(上海人工智能实验室,中国)
;
Sun Yat-Sen University, Shenzhen, China(暨南大学深圳校区,中国)
Fusion or Confusion? Multimodal Complexity Is Not All You Need
融合还是混淆?多模态复杂性并不都是你需要的
Tillmann Rheude, Roland Eils, Benjamin Wild
机构
*
Berlin Institute of Health, Charité - Universitätsmedizin Berlin(柏林健康研究所,柏林查理医院)
;
Intelligent Medicine Institute, Fudan University(复旦大学智能医学研究院)
;
Department of Mathematics and Computer Science, Freie Universität Berlin(柏林自由大学数学与计算机科学系)
机构
*
Berlin Institute of Health, Charité - Universitätsmedizin Berlin(柏林健康研究所,柏林查理大学)
;
Intelligent Medicine Institute, Fudan University(智能医学研究院,复旦大学)
;
Department of Mathematics and Computer Science, Freie Universität Berlin(数学与计算机科学系,柏林自由大学)