Masked Generative Transformer Is What You Need for Image Editing
你需要Masked Generative Transformer来进行图像编辑
Wei Chow, Linfeng Li, Xian Sun, Lingdong Kong, Zefeng Li, Qi Xu, Hang Song, Tian Ye, Xian Wang, Jinbin Bai, Shilin Xu, Xiangtai Li, Junting Pan, Shaoteng Liu, Ran Zhou, Tianshu Yang, Songhua Liu
机构
*
ByteDance(字节跳动)
;
National University of Singapore(新加坡国立大学)
;
Duke University(杜克大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
HKUST(GZ)(香港科技大学(广州))
机构
*
University of Louisville(路易斯维尔大学)
;
National University of Singapore(新加坡国立大学)
;
Florida International University(佛罗里达国际大学)
;
Shenyang University of Chemical Technology(沈阳化学工业大学)
;
Huazhong Agricultural University(华中农业大学)
C-CoT: Counterfactual Chain-of-Thought with Vision-Language Models for Safe Autonomous Driving
C-CoT:基于视觉-语言模型的反事实链式推理用于安全自动驾驶
Kefei Tian, Yuansheng Lian, Kai Yang, Xiangdong Chen, Shen Li
机构
*
College of Transportation, Tongji University(同济大学交通运输学院)
;
Department of Civil Engineering, Tsinghua University(清华大学土木工程系)
;
School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动系统学院)
;
Department of Civil and Environmental Engineering, National University of Singapore(新加坡国立大学土木与环境工程系)
FORGE: Fragment-Oriented Ranking and Generation for Context-Aware Molecular Optimization
FORGE:面向上下文的分子优化片段排序与生成
Qingchuan Zhang, He Cao, Hao Li, Yanjun Shao, Zhiyuan Liu, Shihang Wang, Shufang Xie, Shenghua Gao, Xinwu Ye
机构
*
University of Science and Technology of China(中国科学技术大学)
;
International Digital Economy Academy(国际数字经济学院)
;
Peking University(北京大学)
;
Yale University(耶鲁大学)
;
National University of Singapore(新加坡国立大学)
;
Macao Polytechnic University(澳门理工学院)
;
Zhongguancun Academy(中关村学院)
;
University of Hong Kong(香港大学)
VPD-100K: Towards Generalizable and Fine-grained Visual Privacy Protection
VPD-100K: 向通用化和细粒度的视觉隐私保护迈进
Xiaobin Hu, Enpu Zuo, Lanping Hu, Kaiwen Yang, Dianshu Liao, Tianyi Zhang, Bo Yin, Yinsi Zhou, Shidong Pan, Xiaoyu Sun
机构
*
National University of Singapore(新加坡国立大学)
;
Australian National University(澳大利亚国立大学)
;
New York University(纽约大学)
;
The University of New South Wales(新南威尔士大学)
CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Video Subtitle Removal
CLEAR: 基于上下文的端到端无掩码推理的自适应视频字幕移除
Qingdong He, Chaoyi Wang, Peng Tang, Yifan Yang, Xiaobin Hu
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Technical University of Munich(慕尼黑技术大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
National University of Singapore(新加坡国立大学)
机构
*
Saw Swee Hock School of Public Health, National University of Singapore(新加坡国立大学 Saw Swee Hock 公共卫生学院)
;
Institute of Data Science, National University of Singapore(新加坡国立大学数据科学研究所)
;
Guangzhou Research Translation and Innovation Institute, National University of Singapore(新加坡国立大学广州研究翻译与创新研究所)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Beijing Key Laboratory of Brainnetome and Brain-Computer Interface, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所脑网络与脑机接口重点实验室)
;
Brainnetome Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所脑网络中心)
机构
*
Department of Computer Science, School of Computing, National University of Singapore(新加坡国立大学计算机科学系)
;
School of Civil Engineering, Faculty of Engineering, The University of Sydney(悉尼大学土木工程学院)
;
Delft Institute of Applied Mathematics, Delft University of Technology(代尔夫特理工大学应用数学研究所)
Alignment-Sensitive Minimax Rates for Spectral Algorithms with Learned Kernels
对具有学习核的谱算法的对齐敏感最小最大率
Dongming Huang, Zhifan Li, Yicheng Li, Qian Lin
机构
*
Department of Statistics and Data Science, National University of Singapore, Singapore(新加坡国立大学统计与数据科学系)
;
School of Statistics and Mathematics, Zhongnan University of Economics and Law, Wuhan, China(中南财经政法大学统计与数学学院)
;
Department of Statistics and Data Science, Tsinghua University, Beijing, China(清华大学统计与数据科学系)
机构
*
Saw Swee Hock School of Public Health, National University of Singapore(新加坡国立大学 Saw Swee Hock 公共卫生学院)
;
Institute of Data Science, National University of Singapore(新加坡国立大学数据科学研究所)
;
Guangzhou Research Translation and Innovation Institute, National University of Singapore(新加坡国立大学广州研究翻译与创新研究所)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Beijing Key Laboratory of Brainnetome and Brain-Computer Interface, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所脑网络与脑机接口重点实验室)
;
Brainnetome Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所脑网络中心)
机构
*
Faculty of Computer Science and Technology, Qilu University of Technology (Shandong Academy of Sciences)(计算机科学与技术学院,齐鲁工业大学(山东省科学院))
;
School of Computing, National University of Singapore(国立新加坡大学计算机学院)
;
Shandong Artificial Intelligence Institute, Qilu University of Technology (Shandong Academy of Sciences)(山东省人工智能研究院,齐鲁工业大学(山东省科学院))
;
Key Laboratory of Computer Vision and System, Ministry of Education, Tianjin University of Technology(教育部计算机视觉与系统重点实验室,天津工业大学)
机构
*
Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系)
;
Department of Bioengineering, University of Pennsylvania(宾夕法尼亚大学生物工程系)
;
Department of Computer and Information Science, University of Pennsylvania(宾夕法尼亚大学计算机与信息科学系)
;
Centre for Computational Biology, Duke-NUS Medical School(杜克-新加坡国立大学医学学校计算生物学中心)
On Uniform Error Bounds for Kernel Regression under Non-Gaussian Noise
关于在非高斯噪声下核回归的统一误差界
Johannes Teutsch, Oleksii Molodchyk, Marion Leibold, Timm Faulwasser, Armin Lederer
机构
*
Chair of Automatic Control Engineering, Department of Computer Engineering, Technical University of Munich(自动控制工程学系,计算机工程系,慕尼黑技术大学)
;
Institute of Control Systems, Hamburg University of Technology(控制系统研究所,汉堡技术大学)
;
Department of Electrical and Computer Engineering, National University of Singapore(电子与计算机工程系,新加坡国立大学)
Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities
荒诞世界:一种简单却强大的方法,用于将现实世界扭曲以探测LLM推理能力
Ryan Albright, Golam Md Muktadir, Zarif Ikram, S M Jubaer, Mehrab Hossain, Dianbo Liu
机构
*
The Nueva School(新维学校)
;
University of Southern California(南加州大学)
;
Notre Dame College(诺特大学)
;
Arizona State University(亚利桑那州立大学)
;
National University of Singapore(新加坡国立大学)
PDEAgent-Bench: A Multi-Metric, Multi-Library Benchmark for PDE Solver Generation
PDEAgent-Bench: 一个多指标、多库的PDE求解器生成基准
Zhen Hang, Yushan Yashengjiang, Junhui Li, Huanshuo Dong, Yang Wei, Zhezheng Hao, Jiangtao Ma, Songlin Bai, Haozhong Kai, Xihang Yue, Gangzong Si, Dongming Jiang, Chao Yao, Zhanhua Hu, Jiangqing Zhang, Pengwei Liu, Yaomin Shen, Xingyu Ren, Lei Liu, Zikang Xu, Han Li, Qingsong Yao, Hande Dong, Hong Wang
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Tencent(腾讯)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Zhejiang University(浙江大学)
;
National University of Singapore(新加坡国立大学)
;
Tsinghua University(清华大学)
;
University of Texas at Dallas(德克萨斯大学达拉斯分校)
;
Arizona State University(亚利桑那州立大学)
;
Rice University(里士满大学)
;
Technical University of Munich(慕尼黑技术大学)
;
Stanford University(斯坦福大学)
;
Alibaba Group(阿里巴巴集团)
CLR-voyance: Reinforcing Open-Ended Reasoning for Inpatient Clinical Decision Support with Outcome-Aware Rubrics
CLR-voyance:通过结果感知的评分表强化住院患者临床决策支持中的开放性推理
Aishik Nagar, Arun-Kumar Kaliya-Perumal, Yu-Hsuan Han, Andrew Sheng-Han Huang, Kristen Kee, Yushi Cao, Yiming Chen, Hongchao Jiang
机构
*
ASUS Intelligent Cloud Services (AICS)(ASUS智能云服务(AICS))
;
Rehabilitation Research Institute of Singapore, Nanyang Technological University(新加坡康复研究院,南洋理工大学)
;
Department of Family Medicine, Taipei Veterans General Hospital(台北荣民总医院家庭医学部)
;
School of Medicine, National Yang Ming Chiao Tung University(国家阳明交通大学医学院)
;
Yong Loo Lin School of Medicine, National University of Singapore(新加坡国立大学 Yong Loo Lin 医学院)
机构
*
Carnegie Mellon University(卡内基梅隆大学)
;
University of Science and Technology of China(中国科学技术大学)
;
National University of Singapore(新加坡国立大学)
;
Shanghai AI Lab(上海人工智能实验室)
Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs
无参与的感知:LMMs中因果发现缺陷的剖析
Jiafeng Liang, Zhihao Zhu, Zihan Zhang, Baoqi Ren, Shixin Jiang, Runxuan Liu, Tao Ren, Ming Liu, See-Kiong Ng, Bing Qin
机构
*
Harbin Institute of Technology(哈尔滨工业大学)
;
Pengcheng Laboratory(鹏城实验室)
;
National University of Singapore(新加坡国立大学)
;
Peking University(北京大学)
;
Harvard University(哈佛大学)
机构
*
University of the Chinese Academy of Sciences(中国科学院大学)
;
National University of Singapore(新加坡国立大学)
;
Zhejiang University(浙江大学)
;
State Key Laboratory of Communication Content Cognition, People’s Daily Online(人民日報網通信內容認知重點實驗室)