MM-ACT: Learn from Multimodal Parallel Generation to Act
MM-ACT: 从多模态并行生成中学习以行动
Haotian Liang, Xinyi Chen, Bin Wang, Mingkang Chen, Yitian Liu, Yuhao Zhang, Zanxin Chen, Tianshuo Yang, Yilun Chen, Jiangmiao Pang, Dong Liu, Xiaokang Yang, Yao Mu, Wenqi Shao, Ping Luo
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
The University of Hong Kong(香港大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Fudan University(复旦大学)
;
Zhejiang University(浙江大学)
Towards Unsupervised Domain Bridging via Image Degradation in Semantic Segmentation
通过图像退化实现无监督领域桥接的语义分割方法
Wangkai Li, Rui Sun, Huayu Mai, Tianzhu Zhang
机构
*
MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知联合实验室,中国科学技术大学)
;
National Key Laboratory of Deep Space Exploration, Deep Space Exploration Laboratory(国家深空探测重点实验室,深空探测实验室)
专题命中
可控生成
:diffusion(abstract);分类 cs.CV
AI总结
DiDA通过图像退化构建中间领域并补偿语义偏移,提升语义分割在不同领域间的适应性能。
CommentsAccepted by Conference on Neural Information Processing Systems (NeurIPS 2025)
机构
*
the Department of Electronic and Computer Engineering, the Hong Kong University of Science and Technology, Hong Kong, China(电子与计算机工程系,香港科技大学,香港,中国)
;
the Academy for Engineering and Technology, Fudan University, Shanghai, China(工程与技术学院,复旦大学,上海,中国)
;
Zhuoyu Technology Co., Ltd., Shenzhen, China(珠海优创科技有限公司,深圳,中国)
HalluGen: Synthesizing Realistic and Controllable Hallucinations for Evaluating Image Restoration
HalluGen: 合成逼真且可控的幻觉以评估图像修复
Seunghoi Kim, Henry F. J. Tregidgo, Chen Jin, Matteo Figini, Daniel C. Alexander
机构
*
Hawkes Institute, UCL(UCL哈维斯研究所)
;
Dept. of Medical Physics and Biomedical Engineering, UCL(UCL医学物理与生物医学工程系)
;
Dept. of Computer Science, UCL(UCL计算机科学系)
;
Centre for AI, DS&AI, AstraZeneca, UK(阿斯利康AI中心)
机构
*
Technical University of Munich(慕尼黑技术大学)
;
Fudan University(复旦大学)
;
National University of Singapore(新加坡国立大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
CzechLynx: A Dataset for Individual Identification and Pose Estimation of the Eurasian Lynx
CzechLynx:用于欧亚猞猁个体识别和姿态估计的数据集
Lukas Picek, Elisa Belotti, Michal Bojda, Ludek Bufka, Vojtech Cermak, Martin Dula, Rostislav Dvorak, Luboslav Hrdy, Miroslav Jirik, Vaclav Kocourek, Josefa Krausova, Jirı Labuda, Jakub Straka, Ludek Toman, Vlado Trulık, Martin Vana, Miroslav Kutal
机构
*
Faculty of Applied Sciences, University of West Bohemia in Pilsen, Czechia(西波西米亚大学应用科学学院)
;
Inria, LIRMM, University of Montpellier, France(Inria、LIRMM、蒙彼利埃大学,法国)
;
Faculty of Forestry and Wood Sciences, Czech University of Life Sciences Prague, Czechia(林业与木科学学院,捷克生命科学大学布拉格,捷克)
;
Department of Research and Nature Protection, Šumava National Park Administration, Czechia(研究与自然保护部门,Šumava国家公园管理局,捷克)
;
Department of Forest Ecology, Faculty of Forestry and Wood Technology, Mendel University in Brno, Czechia(森林生态学系,林业与木技术学院,梅德利大学布拉格,捷克)
;
Friends of the Earth Czech Republic, Carnivore Conservation Programme, Czechia(捷克地球之友、食肉动物保护计划,捷克)
;
Center for Machine Perception, Czech Technical University in Prague, Czechia(机器感知中心,布拉格捷克技术大学,捷克)
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
University of Health and Rehabilitation Sciences(康复科学大学)
;
The University of Aberdeen(阿伯丁大学)
Dream4D: Lifting Camera-Controlled I2V towards Spatiotemporally Consistent 4D Generation
Dream4D: 通过可控视频生成与神经4D重建提升I2V向时空一致4D生成
Xiaoyan Liu, Kangrui Li, Yuehao Song, Jiaxin Liu
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
The University of New South Wales(新南威尔士大学)
PG-ControlNet: A Physics-Guided ControlNet for Generative Spatially Varying Image Deblurring
PG-ControlNet:一种用于生成空间变化图像去模糊的物理引导ControlNet
Hakki Motorcu, Mujdat Cetin
机构
*
1 Computer Science Department, University of Rochester
;
2 Goergen Institute for Data Science \& AIS, University of Rochester
;
Computer Engineering Department, University of Rochester