机构
*
Department of Electronic Engineering, Tsinghua University(清华大学电子工程系)
;
Key Laboratory of Marine Robotics, Shenyang(沈阳海洋机器人重点实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(香港科技大学电子及计算机工程学系)
机构
*
The Hong Kong University of Science and Technology(香港科技大学)
;
Hong Kong Generative AI Research and Development Center (HKGAI)(香港生成式人工智能研发中心(HKGAI))
SAIL: Scene-aware Adaptive Iterative Learning for Long-Tail Trajectory Prediction in Autonomous Vehicles
SAIL:面向场景的自适应迭代学习用于自动驾驶车辆的长尾轨迹预测
Bin Rao, Haicheng Liao, Chengyue Wang, Keqiang Li, Zhenning Li, Hai Yang
机构
*
State Key Laboratory of Internet of Things for Smart City and Department of Civil and Environmental Engineering, University of Macau(澳门大学智慧城市物联网国家重点实验室及土木与环境工程系)
;
State Key Laboratory of Internet of Things for Smart City and Department of Computer and Information Science, University of Macau(澳门大学智慧城市物联网国家重点实验室及计算机与信息科学系)
;
Department of Automotive Engineering, Tsinghua University(清华大学汽车工程系)
;
State Key Laboratory of Internet of Things for Smart City and Departments of Civil and Environmental Engineering and Computer and Information Science, University of Macau(澳门大学智慧城市物联网国家重点实验室及土木与环境工程系和计算机与信息科学系)
;
Department of Civil and Environmental Engineering, The Hong Kong University of Science and Technology(香港科技大学土木与环境工程系)
机构
*
Australian Artificial Intelligence Institute, University of Technology Sydney(澳大利亚人工智能研究所,悉尼科技大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Shenzhen University(深圳大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Sichuan University(四川大学)
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
The University of Hong Kong(香港大学)
;
Stony Brook University(石溪大学)
;
Nanyang Technological University(南洋理工大学)
;
Datawhale Org.(Datawhale 组织)
;
Independent Researcher(独立研究者)
Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation
重新思考位置嵌入作为多参考和多镜头视频生成的上下文控制器
Binyuan Huang, Yuning Lu, Weinan Jia, Hualiang Wang, Mu Liu, Daiqing Yang
机构
*
Wuhan University(武汉大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Tsinghua University(清华大学)
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
The Hong Kong University of Science and Technology(香港科技大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Shanghai Jiao Tong University(上海交通大学)
TABQAWORLD: Optimizing Multimodal Reasoning for Multi-Turn Table Question Answering
TABQAWORLD: 优化多模态推理以实现多轮表格问答
Tung Sum Thomas Kwok, Xinyu Wang, Xiaofeng Lin, Peng Lu, Chunhe Wang, Changlun Li, Hanwei Wu, Nan Tang, Elisa Kreiss, Guang Cheng
机构
*
University of California, Los Angeles(加州大学洛杉矶分校)
;
McGill University(麦吉尔大学)
;
Université de Montréal(蒙特利尔大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
SimpleWay
ArchMap: Arch-Flattening and Knowledge-Guided Vision Language Model for Tooth Counting and Structured Dental Understanding
ArchMap:用于牙齿计数和结构牙科理解的拱形扁平化和知识引导的视觉语言模型
Bohan Zhang, Yiyi Miao, Taoyu Wu, Tong Chen, Ji Jiang, Zhuoxiao Li, Zhe Tang, Limin Yu, Jionglong Su
机构
*
Xi'an Jiaotong-Liverpool University(西交利物浦大学)
;
University of Liverpool(利物浦大学)
;
Zhejiang University of Technology(浙江工业大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))