INTACT: Isomorphic Intent-to-Action Learning for Search-Free World Models
INTACT:用于无搜索世界模型的同构意图到动作学习
Junhan Sun, Hao Zhao, Guofeng Zhang
机构
*
State Key Laboratory of CAD&CG, Zhejiang University(浙江大学计算机辅助设计与图形学国家重点实验室)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
;
InSpatio
LaP-Forensics: Latent-Pixel Consistency Guided Multimodal Reasoning for Deepfake Detection
LaP-Forensics:用于深度伪造检测的潜在像素一致性引导的多模态推理
Can Wang, Yuhao Wang, Yushe Cao, Canran Xiao, Fei Shen
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
University College London(伦敦大学学院)
;
Tsinghua University(清华大学)
;
Sun Yat-sen University(中山大学)
;
National University of Singapore(新加坡国立大学)
机构
*
Tsinghua University(清华大学)
;
Qwen Business Unit of Alibaba(阿里巴巴的通义业务部)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Peking University(北京大学)
Retraction-Free Optimization over the Stiefel Manifold for the LoRA Fine-Tuning
用于LoRA微调的Stiefel流形上无回缩优化
Yuan Zhang, Jiang Hu, Zhijian Lai, Lin Lin, Zaiwen Wen
机构
*
Center for Data Science, Peking University(北京大学数据科学中心)
;
Yau Mathematical Sciences Center, Tsinghua University(清华大学丘成桐数学科学中心)
;
Beijing International Center for Mathematical Research, Peking University(北京大学北京国际数学研究中心)
;
Department of Mathematics, University of California, Berkeley(美国加州大学伯克利分校数学系)
;
Center for Machine Learning Research and Changsha Institute for Computing and Digital Economy, Peking University(北京大学机器学习研究中心和长沙计算与数字经济研究院)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
University of Adelaide(阿德莱德大学)
;
Tsinghua University(清华大学)
;
National University of Singapore(新加坡国立大学)
;
Peking University(北京大学)
Jinpeng Chen, Ziyu Yu, Tao Wang, Jun Ma, Hongbo Gao, Senzhang Wang, Zufeng Zhang, Kaimin Wei
机构
*
School of Computer Science (National Pilot Software Engineering School), Beijing University of Posts and Telecommunications(北京邮电大学计算机学院(国家示范性软件学院))
;
Department of Automation, School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学技术学院自动化系)
;
School of Computer Science and Engineering, Central South University(中南大学计算机科学与工程学院)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
College of Information Science and Technology, Jinan University(暨南大学信息科学技术学院)
机构
*
College of Geodesy and Geomatics, Shandong University of Science and Technology(山东科技大学测绘与地理信息学院)
;
School of Environmental Science and Spatial Informatics, China University of Mining and Technology(中国矿业大学环境与测绘学院)
;
State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing, Wuhan University(武汉大学测绘遥感信息工程国家重点实验室)
;
Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)
;
School of Automation, Southeast University(东南大学自动化学院)
;
Department of Geography, National University of Singapore(新加坡国立大学地理系)
ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding
ClinFusion:用于整体医学理解的以视觉为中心的多模态大语言模型系统
Hangjie Yuan, Yichen Qian, Zhiwei Tang, Xianzhe Xu, Lirong Wu, Sicheng Yang, Jinwang Wang, Pengju Wang, Zhitao Zeng, Yizeng Han, Yan Xing, Shengxuan Luo, Tao Feng, Qing Xie, Weigen Yao, Yi Yang, Zuozhu Liu, Jiasheng Tang, Shaocheng Wang, Jitao Wang, Jiahong Dong, Weihua Chen, Feng Xu, Fan Wang
机构
*
DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团)
;
Hupan Laboratory(湖畔实验室)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
Department of Radiology, The Affiliated Yangming Hospital of Ningbo University(宁波大学附属阳明医院放射科)
;
Zhejiang University-University of Illinois Urbana-Champaign Institute, Zhejiang University(浙江大学伊利诺伊大学厄巴纳香槟校区联合学院,浙江大学)
;
Hepato-Pancreato-Biliary Center, Beijing Tsinghua Changgung Hospital, School of Clinical Medicine, Tsinghua Medicine, Tsinghua University(清华长庚医院肝胆胰中心,清华大学临床医学院,清华医学,清华大学)
;
School of Software, Tsinghua University(清华大学软件学院)
;
Beijing National Research Center for Information Science and Technology, Tsinghua University(清华大学北京信息科学与技术国家研究中心)
Why Does Grounding Hurt Medical VQA? Benchmarking, Diagnosis, and Fine-Tuning of Vision-Language Models
对前沿视觉-语言模型进行审计以实现可信的医学视觉问答:定位失败、格式崩溃和领域适应
Xupeng Chen, Binbin Shi, Chenqian Le, Qifu Yin, Lang Lin, Haowei Ni, Ran Gong, Panfeng Li
机构
*
New York University, New York, USA(纽约大学)
;
Tsinghua University, Beijing, China(清华大学)
;
Columbia University, New York, USA(哥伦比亚大学)
;
University of Michigan, Ann Arbor, USA(密歇根大学)
机构
*
Wizard Intelligence Learning Lab, Stanford University(智能巫师学习实验室,斯坦福大学)
;
Peking University(北京大学)
;
Tsinghua University(清华大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
Emotion Collider: Dual Hyperbolic Mirror Manifolds for Sentiment Recovery via Anti Emotion Reflection
情感碰撞:双双曲镜面流形用于通过反情感反射进行情感恢复
Rong Fu, Ziming Wang, Shuo Yin, Kun Liu, Xianda Li, Simon Fong
机构
*
University of Macau(澳门大学)
;
Zhejiang University(浙江大学)
;
Tsinghua University(清华大学)
;
Tongji University(同济大学)
;
University of Southampton(南安普顿大学)
;
University of Bologna(博洛尼亚大学)
;
Minzu University of China(中央民族大学)
CityGuard: Graph-Aware Private Descriptors for Bias-Resilient Identity Search Across Urban Cameras
CityGuard:图感知的隐私保护描述符用于城市摄像头间鲁棒的身份搜索
Rong Fu, Yibo Meng, Jia Yee Tan, Rui Lu, Jiekai Wu, Simon Fong
机构
*
University of Macau(澳门大学)
;
Tsinghua University(清华大学)
;
Renmin University of China(中国人民大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Juntendo University(顺天堂大学)
;
Peking University(北京大学)
LLM-generated personalized nudges for improving pro-environmental behavior: Field evidence from resource conservation
利用大型语言模型的迭代个性化增强行为提示:一项关于电力和热水节约的实地实验
Zonghan Li, Yi Liu, Chunyan Wang, Song Tong, Kaiping Peng, Feng Ji
机构
*
School of Environment, Tsinghua University(清华大学环境学院)
;
Department of Applied Psychology and Human Development, University of Toronto(多伦多大学应用心理学与人类发展系)
;
State Key Laboratory of Regional Environment and Sustainability, Tsinghua University(清华大学区域环境与可持续性国家重点实验室)
;
Department of Psychology, Beijing Normal University at Zhuhai(北京师范大学珠海校区心理学系)
;
Department of Psychology and Cognitive Sciences, Tsinghua University(清华大学心理学与认知科学系)
VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation
VisRAG2.0:通过视觉检索增强生成中的证据引导多图像推理减轻视觉幻觉
Yubo Sun, Chunyi Peng, Yukun Yan, Shi Yu, Zhenghao Liu, Sen Mei, Chi Chen, Maosong Sun
机构
*
School of Software and Microelectronics, Peking University, China(北京大学软件与微电子学院)
;
School of Computer Science and Engineering, Northeastern University, China(东北大学计算机科学与工程学院)
;
Department of Computer Science and Technology, Institute for AI, Tsinghua University, China(清华大学人工智能研究院计算机科学与技术系)