机构
*
Generative AI Lab, College of Computing and Data Science, Nanyang Technological University, Singapore(生成式人工智能实验室,计算与数据科学学院,南洋理工大学,新加坡)
;
Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团)
RESCHED: Rethinking Flexible Job Shop Scheduling from a Transformer-based Architecture with Simplified States
RESCHED: 从基于Transformer的架构重新思考柔性作业车间调度
Xiangjie Xiao, Cong Zhang, Wen Song, Zhiguang Cao
机构
*
School of Computing and Information Systems, Singapore Management University, Singapore(新加坡管理大学计算机与信息系)
;
College of Computing and Data Science, Nanyang Technological University, Singapore(南洋理工大学计算与数据科学学院)
;
Institute of Marine Science and Technology, Shandong University, China(山东大学海洋科学与技术研究院)
SurgCUT3R: Surgical Scene-Aware Continuous Understanding of Temporal 3D Representation
SurgCUT3R: 术场景感知的连续时间3D表示理解
Kaiyuan Xu, Fangzhou Hong, Daniel Elson, Baoru Huang
机构
*
The Hamlyn Centre for Robotic Surgery, Imperial College London(机器人手术中心,帝国理工学院伦敦分校)
;
S-Lab, College of Computing and Data Science, Nanyang Technological University(S实验室, computing and Data Science学院,南洋理工大学)
;
Department of Computer Science, University of Liverpool(计算机科学系,利物浦大学)
StreamVoiceAnon+: Emotion-Preserving Streaming Speaker Anonymization via Frame-Level Acoustic Distillation
StreamVoiceAnon+: 通过帧级声学蒸馏实现情感保留的流式语音匿名化
Nikita Kuzmin, Kong Aik Lee, Eng Siong Chng
机构
*
Nanyang Technological University, Singapore(南洋理工大学)
;
Institute for Infocomm Research (I2R), A*STAR, Singapore(信息与通信研究所)
;
The Hong Kong Polytechnic University, Hong Kong(香港理工大学)
机构
*
Tianjin Key Laboratory of Cognitive Computing and Application(天津认知计算与应用重点实验室)
;
College of Intelligence and Computing(智能与计算学院)
;
Tianjin University(天津大学)
;
TeleAI, China Telecom(TeleAI,中国电信)
;
Southeast University(东南大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Huiyan Technology Co., Ltd(慧言科技有限公司)
;
Nanyang Technological University(南洋理工大学)
;
Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究所,中国科学院)
AuthFace: Towards Authentic Blind Face Restoration with Face-oriented Generative Diffusion Prior
AuthFace: 向真实盲脸修复迈进的面向面部生成扩散先验
Guoqiang Liang, Qingnan Fan, Bingtao Fu, Jinwei Chen, Hong Gu, Lin Wang
机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
vivo Mobile Communication Co., Ltd(vivo移动通信有限公司)
;
Nanyang Technological University(南洋理工大学)
Physics-consistent deep learning for blind aberration recovery in mobile optics
基于物理的深度学习用于移动光学中盲性像差恢复
Kartik Jhawar, Tamo Sancho Miguel Tandoc, Khoo Jun Xuan, Wang Lipo
机构
*
Institute for Digital Molecular Analytics and Science (IDMxS), Nanyang Technological University, Singapore 636921(数字分子分析与科学研究所(IDMxS)、南洋理工大学,新加坡636921)
;
School of Electrical and Electronic Engineering, Nanyang Technological University, Singapore 639798(电气与电子工程学院、南洋理工大学,新加坡639798)
Sicheng Li, Zaiwang Gu, Jie Zhang, Qing Guo, Xudong Jiang, Jun Cheng
机构
*
School of Electrical and Electronic Engineering, Nanyang Technological University (NTU), Singapore(南洋理工大学电子与电气工程学院)
;
Institute for Infocomm Research (I 2 R), Agency for Science, Technology and Research (A*STAR), Singapore(信息与通信研究 institute(I2R),科技研究局(A*STAR))
;
Institute of High Performance Computing, Agency for Science, Technology and Research (A*STAR), Singapore(高性能计算研究所,科技研究局(A*STAR))
;
College of Computer Science, Nankai University, Tianjin, China(南开大学计算机学院)
机构
*
Department of Artificial Intelligence, Xi'an Jiaotong University(人工智能系,西安交通大学)
;
College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学)
;
School of Computer Science and Technology, Harbin Institute of Technology(计算机科学与技术学院,哈尔滨工业大学)
AI总结
逐步细化调节通过动态控制细化规则提升扩散语言模型解码效率并保持生成质量。
Comments19 pages, 10 figures, Code available upon publication
NOVA3R: Non-pixel-aligned Visual Transformer for Amodal 3D Reconstruction
NOVA3R:非像素对齐的视觉变换器用于模态3D重建
Weirong Chen, Chuanxia Zheng, Ganlin Zhang, Andrea Vedaldi, Daniel Cremers
机构
*
Technical University of Munich(慕尼黑技术大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
University of Oxford(牛津大学)
;
Nanyang Technological University(南洋理工大学)
机构
*
Tianjin Key Laboratory of Intelligent Unmanned Swarm Technology and System, School of Electrical and Information Engineering, Tianjin University(天津智能无人群技术与系统重点实验室,电气与信息工程学院,天津大学)
;
School of Mechanical and Aerospace Engineering, Nanyang Technological University(机械与航空航天工程学院,南洋理工大学)
;
Shenzhen Institute for Advanced Study, University of Electronic Science and Technology of China(深圳高等研究院,电子科学与技术大学)
机构
*
AGI lab, Westlake University(西溪大学AGI实验室)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Nanyang Technological University(南洋理工大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Science and Technology of China(中国科学技术大学)
MambaTAD: When State-Space Models Meet Long-Range Temporal Action Detection
当状态空间模型与长程时序动作检测相遇:MambaTAD
Hui Lu, Yi Yu, Shijian Lu, Deepu Rajan, Boon Poh Ng, Alex C. Kot, Xudong Jiang
机构
*
Rapid-Rich Object Search Lab, Interdisciplinary Graduate Programme, Nanyang Technological University, Singapore(跨学科研究生项目,南洋理工大学,新加坡)
;
College of Computing and Data Science, Nanyang Technological University, Singapore(计算与数据科学学院,南洋理工大学,新加坡)
;
School of Electrical and Electronic Engineering, Nanyang Technological University, Singapore(电子与电气工程学院,南洋理工大学,新加坡)
MachaGrasp: Morphology-Aware Cross-Embodiment Dexterous Hand Articulation Generation for Grasping
MachaGrasp:基于形态的跨躯体灵巧手关节生成方法用于抓取
Heng Zhang, Kevin Yuchen Ma, Mike Zheng Shou, Weisi Lin, Yan Wu
机构
*
Robotics & Autonomous Systems Division, Institute for Infocomm Research, Agency for Science, Technology and Research (A*STAR-I 2 R)(机器人与自主系统 division,信息通信研究所,科技研究局)
;
College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学)
;
Show Lab, National University of Singapore(Show Lab,国立新加坡大学)