$Δ$VLA: Prior-Guided Vision-Language-Action Models via World Knowledge Variation
$Δ$VLA:通过世界知识变化引导的视觉-语言-动作模型
Yijie Zhu, Jie He, Rui Shao, Kaishen Yuan, Tao Tan, Xiaochen Yuan, Zitong Yu
机构
*
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Great Bay University(大湾大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Macao Polytechnic University(澳门理工学院)
ExGS: Extreme 3D Gaussian Compression with Diffusion Priors
ExGS:基于扩散先验的极端3D高斯压缩
Jiaqi Chen, Xinhao Ji, Yuanyuan Gao, Hao Li, Yuning Gong, Yifei Liu, Dan Xu, Zhihang Zhong, Dingwen Zhang, Xiao Sun
机构
*
Northwestern Polytechnical University(北western工业大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
Hong Kong University of Science and Technology(香港科技大学)
Listening with the Eyes: Benchmarking Egocentric Co-Speech Grounding across Space and Time
用眼睛倾听:跨时空的自体视觉共指基准测试
Weijie Zhou, Xuantang Xiong, Zhenlin Hu, Xiaomeng Zhu, Chaoyang Zhao, Honghui Dong, Zhengyou Zhang, Ming Tang, Jinqiao Wang
机构
*
Beijing Jiaotong University(北京交通大学)
;
Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences (CASIA)(基础模型研究中心、自动化研究所、中国科学院(CASIA))
;
Tencent Robotics X(腾讯机器人X)
;
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology (HKUST)(计算机科学与工程系、香港科学与技术大学(HKUST))
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学深圳学院)
TDM-R1: Reinforcing Few-Step Diffusion Models with Non-Differentiable Reward
TDM-R1: 通过非可微奖励强化少步扩散模型
Yihong Luo, Tianyang Hu, Weijian Luo, Jing Tang
机构
*
Hong Kong University of Science and Technology(香港科技大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
hi-Lab, Xiaohongshu Inc(小红书实验室)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
C$^2$-Explorer: Contiguity-Driven Task Allocation with Connectivity-Aware Task Representation for Decentralized Multi-UAV Exploration
C$^2$-Explorer: 基于连通性的任务分配与连接感知的任务表示用于分布式多无人机探索
Xinlu Yan, Mingjie Zhang, Yuhao Fang, Yanke Sun, Jun Ma, Youmin Gong, Boyu Zhou, Jie Mei
机构
*
School of Intelligence Science and Engineering, Harbin Institute of Technology, Shenzhen, Guangdong, China(哈尔滨工业大学深圳研究院)
;
Robotics and Autonomous Systems Thrust, The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)机器人与自主系统方向)
;
Department of Mechanical and Energy Engineering, Southern University of Science and Technology, Shenzhen, Guangdong, China(南方科技大学机械与能源工程系)
机构
*
Shanghai AI Lab(上海人工智能实验室)
;
Northwestern Polytechnical University(西北工业大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Peking University(北京大学)
;
Nanyang Technological University(南洋理工大学)
;
Beihang University(北京航空航天大学)
;
Sichuan University(四川大学)
;
Tsinghua University(清华大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Fudan University(复旦大学)
;
Hong Kong University of Science and Technology(香港科技大学)
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
多模态解耦与耦合网络用于抗干扰的3D目标检测
Rui Ding, Zhaonian Kuang, Yuzhe Ji, Meng Yang, Xinhu Zheng, Gang Hua
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(人机混合增强智能国家重点实验室,人工智能与机器人研究院,西安交通大学)
;
Intelligent Transportation Thrust of the Systems Hub, The Hong Kong University of Science and Technology (Guangzhou)(系统枢纽智能交通方向,香港科技大学(广州))
;
Multimodal Experiences Research Lab, Dolby Laboratories(多模态体验研究实验室,Dolby实验室)
机构
*
Fudan University(复旦大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Xi'an Jiaotong-Liverpool University(西安交通大学利物浦大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
University of British Columbia(不列颠哥伦比亚大学)
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
面向数据高效的单目3D物体检测的物体-场景-相机分解与重组
Zhaonian Kuang, Rui Ding, Meng Yang, Xinhu Zheng, Gang Hua
机构
*
Institute of Artificial Intelligence and Robotics(人工智能与机器人研究院)
;
Xi'an Jiaotong University(西安交通大学)
;
Intelligent Transportation Thrust of the Systems Hub(系统枢纽智能交通事业部)
;
Hong Kong University of Science and Technology (GZ)(香港科技大学(广州))
;
Amazon Alexa AI(亚马逊Alexa AI)
SwingArena: Competitive Programming Arena for Long-context GitHub Issue Solving
SwingArena: 用于长上下文GitHub问题解决的竞争性编程竞技场
Wendong Xu, Jing Xiong, Chenyang Zhao, Qiujiang Chen, Haoran Wang, Hui Shen, Zhongwei Wan, Jianbo Dai, Taiqiang Wu, He Xiao, Chaofan Tao, Z. Morley Mao, Ying Sheng, Zhijiang Guo, Hongxia Yang, Bei Yu, Lingpeng Kong, Quanquan Gu, Ngai Wong
机构
*
The University of Hong Kong(香港大学)
;
University of California Los Angeles(加州大学洛杉矶分校)
;
Tsinghua University(清华大学)
;
University of Michigan Ann Arbor(密歇根大学安娜堡分校)
;
The Ohio State University(俄亥俄州立大学)
;
University of Edinburgh(爱丁堡大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
LMSYS Org(LMSYS组织)
DeepSparse: A Foundation Model for Sparse-View CBCT Reconstruction
DeepSparse: 一种用于稀疏视图CBCT重建的基础模型
Yiqun Lin, Jixiang Chen, Hualiang Wang, Jiewen Yang, Jiarong Guo, Yi Zhang, Xiaomeng Li
机构
*
Department of Electronic and Computer Engineering, the Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学)
;
School of Cyber Science and Engineering, Sichuan University(网络科学与工程学院,四川大学)
The Exploration of Error Bounds in Classification with Noisy Labels
在噪声标签下分类中误差界限的探索
Haixia Liu, Boxiao Li, Can Yang, Yang Wang
机构
*
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
School of Mathematics and Statistics(数学与统计学学院)
;
Huazhong University of Science and Technology(华中科技大学)
;
The University of Hong Kong(香港大学)
Rethinking Deep Research from the Perspective of Web Content Distribution Matching
从网页内容分布匹配视角重新思考深度研究
Zixuan Yu, Zhenheng Tang, Tongliang Liu, Chengqi Zhang, Xiaowen Chu, Bo Han
机构
*
School of Computer science and engineering, Sun Yat-sen University(计算机科学与工程学院,中山大学)
;
CSE, The Hong Kong University of Science and Technology(计算机科学与工程系,香港科学与技术大学)
;
Sydney AI Centre, The University of Sydney(悉尼人工智能中心,悉尼大学)
;
Department of Data Science and Artificial Intelligence, The Hong Kong Polytechnic University(数据科学与人工智能系,香港理工大学)
;
DSA Thrust, The Hong Kong University of Science and Technology (GuangZhou)(数据科学与技术 thrust,香港科学与技术大学(广州))
;
TMLR Group, Department of Computer Science, Hong Kong Baptist University(TMLR 组,计算机科学系,香港 Baptist 大学)
机构
*
Shanghai Artificial Intelligence Laboratory, OpenDataLab, OpenDataArena(上海人工智能实验室、OpenDataLab、OpenDataArena)
;
Hong Kong University of Science and Technology(香港科技大学)
RoTri-Diff: A Spatial Robot-Object Triadic Interaction-Guided Diffusion Model for Bimanual Manipulation
RoTri-Diff: 一种基于空间机器人-物体三元交互引导的扩散模型用于双臂操作
Zixuan Chen, Nga Teng Chan, Yiwen Hou, Chenrui Tie, Zixuan Liu, Haonan Chen, Junting Chen, Jieqi Shi, Yang Gao, Jing Huo, Lin Shao
机构
*
School of Computer Science, Nanjing University(南京大学计算机科学学院)
;
School of Intelligence Science and Technology, Nanjing University(南京大学智能科学与技术学院)
;
School of Computing, National University of Singapore(新加坡国立大学计算机学院)
;
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系)
机构
*
Huawei Foundation Model Department(华为基础模型部门)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
OralGPT-Plus: Learning to Use Visual Tools via Reinforcement Learning for Panoramic X-ray Analysis
OralGPT-Plus:通过强化学习学习使用视觉工具进行全景X射线分析
Yuxuan Fan, Jing Hao, Hong Chen, Jiahao Bao, Yihua Shao, Yuci Liang, Kuo Feng Hung, Hao Tang
机构
*
The Hong Kong University of Science and Technology (GZ)(香港科学与技术大学)
;
Faculty of Dentistry, The University of Hong Kong(香港大学牙医学院)
;
School of Computer Science, Peking University(北京大学计算机学院)
;
Shanghai Jiao Tong University(上海交通大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院)
StruVis: Enhancing Reasoning-based Text-to-Image Generation via Thinking with Structured Vision
StruVis: 通过结构化视觉进行推理的文本到图像生成增强
Yuanhuiyi Lyu, Kaiyu Lei, Ziqiao Weng, Xu Zheng, Lutao Jiang, Teng Li, Yangfu Li, Ziyuan Huang, Linfeng Zhang, Xuming Hu
机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Ant Group(蚂蚁集团)
;
Shanghai Jiao Tong University(上海交通大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
East China Normal University(华东师范大学)
InnoAds-Composer: Efficient Condition Composition for E-Commerce Poster Generation
InnoAds-Composer: 电商海报生成中的高效条件组合
Yuxin Qin, Ke Cao, Haowei Liu, Ao Ma, Fengheng Li, Honghe Zhu, Zheng Zhang, Run Ling, Wei Feng, Xuanhua He, Zhanjie Zhang, Zhen Guo, Haoyi Bian, Jingjing Lv, Junjie Shen, Ching Law
机构
*
JD.com, Inc.(京东公司)
;
Chongqing University of Posts and Telecommunications(重庆邮电大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Zhejiang University(浙江大学)
Cross-Scale Pansharpening via ScaleFormer and the PanScale Benchmark
跨尺度 pansharpening 通过 ScaleFormer 和 PanScale 数据集
Ke Cao, Xuanhua He, Xueheng Li, Lingting Zhu, Yingying Wang, Ao Ma, Zhanjie Zhang, Man Zhou, Chengjun Xie, Jie Zhang
机构
*
HFIPS, Chinese Academy of Sciences(中国科学院HFIPS)
;
University of Science and Technology of China(中国科学技术大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
The University of Hong Kong(香港大学)
;
Xiamen University(厦门大学)
;
Zhejiang University(浙江大学)