Mema: Memory-Augmented Adapter for Enhanced Vision-Language Understanding
Mema:增强视觉-语言理解的内存增强适配器
Ying Liu, Yudong Han, Kean Shi, Liyuan Pan
机构
*
Beijing Institute of Technology(北京理工大学)
;
Peking University(北京大学)
;
Yangtze Delta Region Academy of Beijing Institude of Technology(北京理工大学长江三角洲地区研究院)
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学 赛洛岭人工智能学院)
;
Amap, Alibaba Group(阿里巴巴集团 阿里地图)
;
Peking University(北京大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Southern University of Science and Technology(南方科技大学)
Mobile GUI Agents under Real-world Threats: Are We There Yet?
移动GUI代理在现实威胁下的表现:我们已经到达了吗?
Guohong Liu, Jialei Ye, Jiacheng Liu, Yuanchun Li, Wei Liu, Pengzhi Gao, Jian Luan, Yunxin Liu
机构
*
Institute for AI Industry Research (AIR), Tsinghua University(人工智能产业研究院(AIR),清华大学)
;
University of Electronic Science and Technology of China(电子科学与技术大学)
;
Peking University(北京大学)
;
MiLM Plus, Xiaomi Inc.(MiLM Plus,小米公司)
机构
*
School of Mathematical Sciences, Peking University(北京大学数学科学学院)
;
Center for Statistical Science, Peking University(北京大学统计科学中心)
;
Academy of Mathematics and Systems Science, Chinese Academy of Sciences(中国科学院数学与系统科学研究院)
;
Huawei Foundation Model Dept(华为基金会模型部门)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
机构
*
Basic Model Technology Center, WeChat AI, Tencent Inc.(腾讯基本模型技术中心、微信AI、腾讯公司)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室、计算机学院、北京大学)
GTCN-G: A Residual Graph-Temporal Fusion Network for Imbalanced Intrusion Detection
GTCN-G:一种用于不平衡入侵检测的残差图-时间融合网络
Tianxiang Xu, Zhichao Wen, Xinyu Zhao, Qi Hu, Yan Li, Chang Liu
机构
*
Peking University(北京大学)
;
RWTH Aachen University(亚琛工业大学)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Northeastern University(东北大学)
;
Thales Group(泰雷兹集团)
;
Chinese Medical Information and Big Data Association(中国医学信息与大数据协会)
机构
*
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机学院,北京大学)
;
Basic Model Technology Center, WeChat AI, Tencent Inc.(微信AI基础模型技术中心,腾讯公司)
HistLens: Mapping Idea Change across Concepts and Corpora
HistLens:跨概念和语料库的思想变化映射
Yi Jing, Weiyun Qiu, Yihang Peng, Zhifang Sui
机构
*
Department of Computer Science and Technology, Tsinghua University, China(清华大学计算机科学与技术系)
;
School of History, Nanjing University, China(南京大学历史学院)
;
Department of Chinese Language and Literature, Tsinghua University, China(清华大学中国语言文学系)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University, China(北京大学计算机学院多媒体信息处理国家重点实验室)
Learning from Contrasts: Synthesizing Reasoning Paths from Diverse Search Trajectories
从对比中学习:从多样的搜索轨迹中合成推理路径
Peiyang Liu, Zhirui Chen, Xi Wang, Di Liang, Youru Li, Zhi Cai, Wei Ye
机构
*
National Engineering Research Center for Software Engineering, Peking University(北京大学软件工程国家工程研究中心)
;
School of Software and Microelectronics, Peking University(北京大学软件与微电子学院)
;
UCAS-Terminus AI Lab, University of Chinese Academy of Sciences(中国科学院大学UCAS-Terminus人工智能实验室)
;
Tencent Technology(腾讯科技)
;
College of Computer Science, Beijing University of Technology(北京工业大学计算机学院)
机构
*
Peking University(北京大学)
;
JD Logistics(京东物流)
;
Nankai University(南开大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Rutgers University(罗格斯大学)
;
University of Texas at Dallas(德克萨斯大学达拉斯分校)
;
Institut Polytechnique de Paris(巴黎综合理工学院)
机构
*
Peking University(北京大学)
;
Key Laboratory of Data Intelligence and Security(数据智能与安全重点实验室)
;
Alibaba Group(阿里巴巴集团)
;
Key Laboratory of Data Space Technology and System(数据空间技术与系统重点实验室)
If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs
如果一个大语言模型是一个角色,它会知道自己故事吗?评估大语言模型的终身学习
Siqi Fan, Xiusheng Huang, Yiqun Yao, Xuezhi Fang, Kang Liu, Peng Han, Shuo Shang, Aixin Sun, Yequan Wang
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所复杂系统认知与决策智能重点实验室)
;
Nanyang Technological University(南洋理工大学)
;
Peking University(北京大学)
;
Spin Matrix, China(中国Spin Matrix公司)
A Benchmark and Multi-Agent System for Instruction-driven Cinematic Video Compilation
一个用于指令驱动电影视频编译的基准和多智能体系统
Peixuan Zhang, Chang Zhou, Ziyuan Zhang, Hualuo Liu, Chunjie Zhang, Jingqi Liu, Xiaohui Zhou, Xi Chen, Shuchen Weng, Si Li, Boxin Shi
机构
*
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
AI Technology Center, Online Video Business Unit, Tencent PCG(腾讯PCG在线视频事业部AI技术中心)
;
Tsinghua University(清华大学)
;
Beijing Academy of Artificial Intelligence(北京智源人工智能研究院)
;
State Key Lab of Multimedia Info. Processing, School of Computer Science, Peking University(北京大学计算机学院多媒体信息处理国家重点实验室)
;
Nat’l Eng. Research Ctr. of Visual Technology, School of Computer Science, Peking University(北京大学计算机学院国家视觉技术工程研究中心)
;
School of Software and Microelectronics, Peking University(北京大学软件与微电子学院)
ReContraster: Making Your Posters Stand Out with Regional Contrast
ReContraster:利用区域对比让海报脱颖而出
Peixuan Zhang, Zijian Jia, Ziqi Cai, Shuchen Weng, Si Li, Boxin Shi
机构
*
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
State Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(北京大学计算机学院多媒体信息处理国家重点实验室)
;
National Engineering Research Center of Visual Technology, School of Computer Science, Peking University(北京大学计算机学院国家视觉技术工程研究中心)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
机构
*
School of Software, Henan University(河南大学软件学院)
;
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Peking University(北京大学)
机构
*
Nanjing University(南京大学)
;
Microsoft Research Asia(微软亚洲研究院)
;
Peking University(北京大学)
;
Zhejiang University(浙江大学)
;
Wuhan University(武汉大学)
;
Hong Kong University of Science and Technology(香港科技大学)
Dark-EvGS: Event Camera as an Eye for Radiance Field in the Dark
暗光环境下的辐射场重建:事件相机作为视觉的眼睛
Jingqian Wu, Peiqi Duan, Zongqiang Wang, Changwei Wang, Boxin Shi, Edmund Y. Lam
机构
*
Department of Electrical and Electronic Engineering, The University of Hong Kong(香港大学电机电子工程系)
;
State Key Laboratory of Multimedia Information Processing and National Engineering Research Center of Visual Technology, School of Computer Science, Peking University(北京大学计算机学院多媒体信息处理国家重点实验室及视觉技术国家工程研究中心)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Key Laboratory of Computing Power Network and Information Security, Ministry of Education, Shandong Computer Science Center, Qilu University of Technology(齐鲁工业大学(山东省科学院)山东省计算中心(国家超级计算济南中心)计算力网络与信息安全教育部重点实验室)