机构
*
Nanjing University of Aeronautics and Astronautics(南京航空航天大学)
;
Peking University(北京大学)
;
Independent Researcher(独立研究者)
;
Jinling Clinical Medical College College of Artificial Intelligence Nanjing University of Aeronautics and Astronautics(金陵临床医学院人工智能学院南京航空航天大学)
机构
*
School of Automation, Beijing Institute of Technology(自动化学院,北京理工大学)
;
School of Computer Science, Wuhan University(计算机学院,武汉大学)
;
Great Wall Motor(长城汽车)
;
School of Information and Electronic Engineering, Zhejiang University of Science and Technology(信息电子工程学院,浙江理工大学)
CommentsThis paper is an extended version of the authors' work previously presented at the ICRA conference. To appear in IEEE Transactions on Circuits and Systems for Video Technology. DOl: 10.1109/TCSVT.2026.3701706
机构
*
Hefei University of Technology(合肥工业大学)
;
Intelligent Interconnected Systems Laboratory of Anhui Province(安徽省智能互联系统实验室)
;
Jianghuai Advanced Technology Center(江淮前沿技术中心)
;
Anhui Provincial Industry Innovation Center of Humanoid Robots(安徽省人形机器人产业创新中心)
;
Anhui Provincial Key Laboratory of Humanoid Robots(安徽省人形机器人重点实验室)
Disentangled Fine-Grained Prototype Learning for Incomplete Image-Tabular Classification
面向不完整图像-表格分类的解缠细粒度原型学习
Feixiang Zhou, Jianyang Xie, Zhuangzhi Gao, Qinkai Yu, Fu Wang, Yuheng Fan, Jing Li, Zheheng Jiang, Yitian Zhao, Yanda Meng, He Zhao, Gregory Y. H. Lip, Yalin Zheng
机构
*
School of Eye and Vision Sciences, University of Liverpool, U.K.(利物浦大学眼科与视觉科学学院)
;
Department of Cardiovascular and Metabolic Medicine, University of Liverpool, U.K.(利物浦大学心血管与代谢医学系)
;
School of Computer Science, University of Exeter, U.K.(埃克塞特大学计算机科学学院)
;
School of Computer Science and Engineering, South China University of Technology, China(华南理工大学计算机科学与工程学院)
;
School of Computing and Mathematical Sciences, University of Leicester, U.K.(莱斯特大学计算科学与数学科学学院)
;
Ningbo Institute of Industrial Technology, Chinese Academy of Sciences, China(中国科学院宁波工业技术研究所)
;
Bioengineering Program, Biological and Environmental Science and Engineering Division (BESE), King Abdullah University of Science and Technology (KAUST), Saudi Arabia(卡尔斯塔德大学科学与技术学院(KAUST)生物工程项目,沙特阿拉伯)
UnsOcc: 3D Semantic Occupancy Prediction in Unstructured Scene via Rendering Fusion
UnsOcc:非结构化场景下基于渲染融合的3D语义占用预测
Ye Wu, Ruiqi Song, Baiyong Ding, Nanxin Zeng, Junjie Cheng, Yunfeng Ai
机构
*
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Waytous Inc.(Waytous公司)
MUSCLE-NET: Predicted-Multiscale-Aware Network for Pedestrian Trajectory Forecasting
MUSCLE-NET:面向行人轨迹预测的预测多尺度感知网络
Yu Liu, Ming Huang, Xiao Ren, Zhijie Liu, Youfu Li, He Kong
机构
*
Guangdong Provincial Key Laboratory of Fully Actuated System Control Theory and Technology, School of Automation and Intelligent Manufacturing, Southern University of Science and Technology (SUSTech), Shenzhen(广东省全主动系统控制理论与技术重点实验室,自动化与智能制造学院,南方科技大学(SUSTech),深圳)
;
Department of Mechanical Engineering, City University of Hong Kong, Hong Kong SAR, China(香港城市大学机械工程系,香港特别行政区,中国)
Image-Conditioned Instance Prompt Network for Referring Remote Sensing Image Segmentation
图像条件实例提示网络用于遥感图像指代分割
Biaoyu Ren, Qingsheng Wang, Cun Xu, Dingkang Yang, Wenxuan Wang
机构
*
School of Computer Science, Northwestern Polytechnical University, Xi'an, China(西北工业大学计算机科学学院,西安,中国)
;
College of Intelligent Robotics and Advanced Manufacturing, Fudan University, Shanghai, China(复旦大学智能机器人与先进制造学院,上海,中国)
;
Shenzhen Research Institute of Northwestern Polytechnical University, Shenzhen, China(西北工业大学深圳研究院,深圳,中国)
OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond
OScaR:LLMs及更广泛场景中的极压缩KV缓存量化之奥卡姆之刀
Zunhai Su, Rui Yang, Chao Zhang, Yaxiu Liu, Yifan Zhang, Wei Wu, Jing Xiong, Dayou Du, Xialie Zhuang, Yulei Qian, Yuchen Xie, Yik-Chung Wu, Hongxia Yang, Ngai Wong
机构
*
Tsinghua University(清华大学)
;
Meituan LongCat Team(美团LongCat团队)
;
The University of Hong Kong(香港大学)
;
The University of Edinburgh(爱丁堡大学)
;
UCAS(中国科学技术大学)
;
The Hong Kong Polytechnic University(香港理工大学)
Evaluating Cognitive Age Alignment in Interactive AI Agents
评估交互式AI代理的认知年龄对齐
Yifan Shen, Jiawen Zhang, Jian Xu, Junho Kim, Ismini Lourentzou, Xu Cao, Meihuan Huang
机构
*
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Shenzhen Children's Hospital(深圳儿童医院)
;
Peking University(北京大学)
;
Hong Kong Polytechnic University(香港理工大学)
Rethinking Point Clouds as Sequences: A Causal Next-Token Predictive Learning Framework
重新思考点云作为序列:一种因果性下一标记预测学习框架
Yumeng Yao, Jingzhi Dong, Haowen Gu, Tao Chen, Zonghan Wu, Xiaoshui Huang, Yazhou Yao
机构
*
Nanjing University of Science and Technology(南京理工大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Hangzhou City University(杭州城市学院)
;
East China Normal University(华东师范大学)
专题命中
多模态训练与对齐
:multimodal(abstract);multimodal foundation model(abstract);分类 cs.CV
机构
*
South China University of Technology(华南理工大学)
;
Institute for Super Robotics (Huangpu)(机器人研究所(黄埔))
;
Shanghai Jiao Tong University(上海交通大学)
;
Changsha University of Science and Technology(长沙理工大学)
LandSegmenter: Towards a Flexible Foundation Model for Land Use and Land Cover Mapping
LandSegmenter:面向土地利用与覆盖制图的灵活基础模型
Chenying Liu, Wei Huang, Xiao Xiang Zhu
机构
*
Chair of Data Science in Earth Observation, Technical University of Munich(地球观测数据科学教授职位,慕尼黑技术大学)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML))
机构
*
Department of Language Science and Technology, The Hong Kong Polytechnic University(香港理工大学语言科学与技术系)
;
Laboratory of Digital Image and Intelligent Computation, Shanghai Maritime University(上海 Maritime 大学数字图像与智能计算实验室)
Align then Refine: Text-Guided 3D Prostate Lesion Segmentation
对齐后再细化:基于文本的3D前列腺病变分割
Cuiling Sun, Linkai Peng, Adam Murphy, Elif Keles, Hiten D. Patel, Ashley Ross, Frank Miller, Baris Turkbey, Andrea Mia Bejar, Halil Ertugrul Aktas, Gorkem Durak, Ulas Bagci
机构
*
Department of Radiology, Northwestern University, Chicago, USA(放射科,西北大学,芝加哥,美国)
;
Department of Urology, Northwestern University, Chicago, USA(泌尿科,西北大学,芝加哥,美国)
;
Center for Cancer Research, National Cancer Institute, Bethesda, USA(癌症研究中心,国家癌症研究所,贝塞斯达,美国)
Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning
通过强化学习提升多图像接地在MLLMs中的推理能力
Bob Zhang, Haoran Li, Tao Zhang, Jianan Li, Cilin Yan, Xikai Liu, Jiayin Cai, Yanbin Hao
机构
*
Xiaohongshu Inc.(小红书公司)
;
University of Science and Technology of China(中国科学技术大学)
;
Wuhan University(武汉大学)
;
Technical University of Munich(慕尼黑工业大学)
;
Hefei University of Technology(合肥工业大学)