Ziheng Peng, Huiqi Deng, Haoran Jing, Xuankun Rong, Jiahui Han, Xiting Wang, Na Zou, Xia Hu
机构
*
Renmin University of China(中国人民大学)
;
Xi’an Jiaotong University(西安交通大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Wuhan University(武汉大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
SpatialVAM:Spatial-Aware Multi-View Video Diffusion as a Data-Efficient Robot Policy
多视角视频扩散策略:一种3D空间-时间感知的视频动作模型
Peiyan Li, Yixiang Chen, Yuan Xu, Jiabing Yang, Xiangnan Wu, Jun Guo, Nan Sun, Long Qian, Xinghang Li, Xin Xiao, Jing Liu, Nianfeng Liu, Tao Kong, Yan Huang, Liang Wang, Tieniu Tan
机构
*
New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
FiveAges(五时代)
;
Tsinghua University(清华大学)
;
Xi’an Jiaotong University(西安交通大学)
;
Wuhan University(武汉大学)
;
Nanjing University(南京大学)
机构
*
School of Automation and Intelligent Manufacturing, Southern University of Science and Technology, Shenzhen, China(自动化与智能制造学院,南方科技大学,深圳,中国)
;
Guangdong Provincial Key Laboratory of Fully Actuated System Control Theory and Technology, Southern University of Science and Technology, Shenzhen, China(广东省全驱动系统控制理论与技术重点实验室,南方科技大学,深圳,中国)
;
College of Computer Science and Software Engineering, Shenzhen University, Shenzhen, China(计算机科学与软件工程学院,深圳大学,深圳,中国)
;
Department of Computer Science, City University of Hong Kong, Hong Kong SAR, China(计算机科学系,香港城市大学,香港特别行政区,中国)
;
School of Mathematics and Statistics, Xi'an Jiaotong University, Xi'an, China(数学与统计学学院,西安交通大学,西安,中国)
机构
*
Electronic Materials Research Laboratory, Key Laboratory of the Ministry of Education and International Center for Dielectric Research, School of Electronic Science and Engineering, Xi’an Jiaotong University(电子材料研究中心、教育部重点实验室和国际介电研究中心、电子科学与工程学院,西安交通大学)
;
Micius Laboratory, Henan Academy of Sciences(墨子实验室、河南省科学院)
机构
*
School of Computer Science(计算机科学学院)
;
Technology, Xi’an Jiaotong University, Xi’an, China(技术学院,西安交通大学,西安,中国)
;
Department of Transmedia Art, Xi’an Academy of Fine Arts, Xi’an, China(多媒体艺术系,西安美术学院,西安,中国)
;
Department of Oncology, University of Cambridge, Cambridge, U.K.(肿瘤学系,剑桥大学,剑桥,英国)
;
Language Technology Lab, University of Cambridge, Cambridge, U.K.(语言技术实验室,剑桥大学,剑桥,英国)
;
Institute of High Performance Computing, Agency for Science, Technology(高性能计算研究所,科技研究局)
SymphonyGen: 3D Hierarchical Orchestral Generation with Controllable Harmony Skeleton
SymphonyGen:具有可控和声骨架的3D分层交响乐生成
Xuzheng He, Nan Nan, Zhilin Wang, Ziyue Kang, Zhuoru Mo, Ao Li, Yu Pan, Xiaobing Li, Feng Yu, Xiaohong Guan
机构
*
Department of AI Music and Music Information Technology, Central Conservatory of Music(中央音乐学院人工智能音乐与音乐信息科技系)
;
Frontier Institute of Science and Technology, and Interdisciplinary Research Center of Frontier Science and Technology, Xi’an Jiaotong University(西安交通大学前沿科学与技术研究院及交叉科学与技术 interdisciplinary research center)
;
University of Science and Technology of China(中国科学技术大学)
;
Shenzhen University(深圳大学)
RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension
RefBench-PRO:面向感知与推理的指称表达理解基准
Tianyi Gao, Hao Li, Han Fang, Xin Wei, Xiaodong Dong, Hongbo Sun, Ye Yuan, Zhongjiang He, Jinglin Xu, Jingmin Xin, Hao Sun
机构
*
National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(国家人机混合增强智能重点实验室)
;
National Engineering Research Center for Visual Information and Applications(国家视觉信息与应用工程技术研究中心)
;
Institute of Artificial Intelligence and Robotics(人工智能与机器人研究院)
;
Xi’an Jiaotong University(西安交通大学)
;
Institute of Artificial Intelligence (TeleAI)(人工智能研究院(TeleAI))
;
China Telecom(中国电信)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Science and Technology Beijing(北京科技大学)
GLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels
GLI-AL:一个具有统一解剖-病变标签的多模态胶质瘤MRI标签资源
Xingyu Xiang, Shuang Hao, Fan Wang, Jianhua Ma, Chunfeng Lian
机构
*
Key Laboratory of Biomedical Information Engineering of Ministry of Education, School of Life Science and Technology, Xi’an Jiaotong University(教育部生物医学信息工程重点实验室,西安交通大学生命科学与技术学院)
;
School of Mathematics and Statistics, Xi’an Jiaotong University(西安交通大学数学与统计学院)
;
Research Center for Intelligent Medical Equipment and Devices (IMED), Xi’an Jiaotong University(西安交通大学智能医疗设备与器械研究中心)
EmoAgent-R1: Towards Multimodal Emotion Understanding with Reinforcement Learning-based Dynamic Agent Specialization
EmoAgent-R1:基于强化学习的动态智能体专业化实现多模态情感理解
Lihuang Fang, Yuchen Zou, kebing Jin, Jinghui Qin
机构
*
Guangdong University of Technology(广东工业大学)
;
Southern University of Science and Technology(南方科技大学)
;
Xi’an Jiaotong University(西安交通大学)
;
Guizhou Provincial Laboratory of Big Data, State Key Laboratory of Public Big Data, Guizhou University(贵州大学大数据省级重点实验室、公共大数据国家重点实验室)
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(西安交通大学人工智能与机器人研究所人机混合增强智能技术国家重点实验室)
;
Guangdong Institute of Intelligence Science and Technology(广东省智能科学与技术研究院)
;
Institute of Collaborative Innovation, University of Macau(澳门大学协同创新研究院)
;
Beijing Huahang Institute of Radio Measurement(北京华航无线电测量研究所)
Comments36 pages, including appendices. Revised version with updated theoretical analysis, supplementary material, figures and improved table formatting
Can Video Diffusion Models Predict Past Frames? Bidirectional Cycle Consistency for Reversible Interpolation
视频扩散模型能否预测过去帧?双向循环一致性用于可逆插值
Lingyu Liu, Yaxiong Wang, Li Zhu, Zhedong Zheng
机构
*
School of Software Engineering, Xi'an Jiaotong University(西安交通大学软件工程学院)
;
School of Computer and Information Science, Hefei University of Technology(合肥工业大学计算机与信息科学学院)
;
Faculty of Science and Technology, University of Macau(澳门大学科技学院)
Code-MUE: Measuring Code LLMs' Uncertainty through Execution-based Semantic Interaction Graphs
Code-MUE:通过基于执行的语义交互图测量代码语言模型的不确定性
Xiaoning Ren, Yinxing Xue, Lei Ma, Yuheng Huang
机构
*
Xi’an Jiaotong University(西安交通大学)
;
Institute of AI for Industries, Chinese Academy of Sciences(中国科学院人工智能产业研究院)
;
The University of Tokyo(东京大学)
;
University of Alberta(阿尔伯塔大学)
Song Aesthetics Evaluation with Multi-Stem Attention and Hierarchical Uncertainty Modeling
基于多茎注意力和分层不确定性建模的歌曲美学评估
Yishan Lv, Jing Luo, Boyuan Ju, Yang Zhang, Xinda Wu, Bo Yuan, Xinyu Yang
机构
*
School of Computer Science and Technology, Xi’an Jiaotong University, Xi’an, China(计算机科学与技术学院,西安交通大学,西安,中国)
;
Central Media Technology Institute, Huawei(中央媒体技术研究院,华为)
TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation
TouchThinker: 通过大规模数据和动作感知表示将触觉常识推理扩展到开放世界
Kailin Lyu, Di Wu, Pengwei Zhang, Yuhang Zheng, Yingxin Lai, Long Xiao, Kangyi Wu, Pengna Li, Chen Gao, Lianyu Hu, Xiaobin Hu, Jie Hao, Ce Hao, Weihao Yuan, Shuicheng Yan
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
National University of Singapore(新加坡国立大学)
;
Zhongguancun Academy(中关村学院)
;
Xiamen University(厦门大学)
;
Xi’an Jiaotong University(西安交通大学)
;
Nanyang Technological University(南洋理工大学)
;
Nanjing University(南京大学)
机构
*
Research Center for Space Computing System, Zhejiang Lab, Hangzhou(杭州浙大实验室空间计算系统研究中心)
;
Hangzhou Institute for Advanced Study, University of Chinese Academy of Sciences, Hangzhou(中国科学院大学杭州高等研究院)
;
School of Cyber Science and Engineering, Zhengzhou University(郑州大学计算机科学与工程学院)
;
School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院)