CogniVerse: Revolutionizing Multi-Modal Retrieval-Augmented Generation with Cognitive Reflection and Geometric Reasoning
CogniVerse: 用认知反思与几何推理革新多模态检索增强生成
Xiang Fang, Wanlong Fang, Changshuo Wang
机构
*
School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件学院)
;
Nanyang Technological University, Singapore(新加坡南洋理工大学)
;
University College London(伦敦大学学院)
BitC-3DGS: High-Capacity 3D Gaussian Splatting Watermarking via Bit Compression
BitC-3DGS: 基于位压缩的高容量3D高斯泼溅水印技术
Yuquan Bi, Baosheng Yu, Yingke Lei, Jianwei Yang, Hongsong Wang, Jie Gui, Yuan Yan Tang, James Tin-Yau Kwok
机构
*
School of Cyber Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
Lee Kong Chian School of Medicine, Nanyang Technological University(南洋理工大学李科金医学院)
;
College of Electronic Engineer, National University of Defense and Technology(国防科技大学电子工程学院)
;
Institute of AI for Industries, Chinese Academy of Sciences(中国科学院人工智能产业研究所)
;
School of Computer Science and Engineering, Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications, Ministry of Education, Southeast University(东南大学计算机科学与工程学院,教育部新一代人工智能技术及其交叉应用重点实验室)
;
Department of Computer and Information Science, University of Macau(澳门大学计算机与信息科学系)
;
Faculty of Science and Technology, UOW College Hong Kong(UOW学院香港科技学院理学院)
;
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(香港科学与技术大学计算机科学与工程学院)
机构
*
School of Computer Science, Shanghai Jiao Tong University, Shanghai, China(上海交通大学计算机科学学院)
;
Nanyang Technological University, Singapore(南洋理工大学)
;
National University of Singapore, Singapore(新加坡国立大学)
Solved in Unit Domain: JacobiNet for Differentiable Coordinate-Transformed PINNs
在单位域中求解:用于可微坐标变换PINNs的JacobiNet
Xi Chen, Jianchuan Yang, Junjie Zhang, Runnan Yang, Xu Liu, Hong Wang, Tinghui Zheng, Ziyu Ren, Wenqi Hu
机构
*
Department of Mechanical and Aerospace Engineering, The Hong Kong University of Science and Technology(香港科技大学机械与航空航天工程系)
;
Department of Mechanics & Engineering, College Architecture & Environment, Sichuan University(四川大学力学与工程学院)
;
School of Mechanical Engineering and Automation, Beihang University(北京航空航天大学机械工程与自动化学院)
;
Department of Civil and Environmental Engineering, The Hong Kong University of Science and Technology(香港科技大学土木与环境工程系)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算机与数据科学学院)
;
West China Biomedical Big Data Center, West China Hospital, Sichuan University(四川大学西昌生物医学大数据中心,西昌医院)
Mining Multi-Modality Spatio-Temporal Cues for Video Important Person Identification
挖掘多模态时空线索用于视频重要人物识别
Xiao Wang, Minglei Yang, Bin Yang, Wenke Huang, Zheng Wang, Xin Xu, Mang Ye
机构
*
School of Computer Science and Technology, Wuhan University of Science and Technology(武汉科技大学计算机科学与技术学院)
;
Hubei Province Key Laboratory of Intelligent Information Processing and Real-time Industrial System, Wuhan University of Science and Technology(湖北省智能信息处理与实时工业系统重点实验室)
;
School of Computer Science, National Engineering Research Center for Multimedia Software, Hubei Key Laboratory of Multimedia and Network Communication Engineering, Wuhan University(计算机科学学院,国家多媒体软件工程技术研究中心,湖北省多媒体与网络通信工程重点实验室,武汉大学)
;
College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学)
StoryLens: Preference-Aligned Story Rewriting via Context-Aware Narrative Enrichment
StoryLens: 通过上下文感知叙事丰富实现偏好对齐的故事重写
Hanwen Cui, Yuting Mei, Yuhang Fu, Dingyi Yang, Qin Jin
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
AIM3 Lab, Renmin University of China(中国人民大学AIM3实验室)
;
Nanyang Technological University(南洋理工大学)
Structure-Guided Visual Perturbation Neutralization for LVLMs
结构引导的视觉扰动中和用于大型视觉语言模型
Yuanhe Zhang, Xueting Wang, YanBin Ren, Haoran Gao, Xinhan Zheng, Zhenhong Zhou, Fanyu Meng, Li Sun, Sen Su
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
University of Science and Technology of China(中国科学技术大学)
;
JIUTIAN Research(JIUTIAN研究所)
;
Nanyang Technological University(南洋理工大学)
;
Chongqing University of Posts and Telecommunications(重庆邮电大学)
Rethinking Video-Language Model from the Language Input Perspective
从语言输入角度重新思考视频-语言模型
Xiang Fang, Wanlong Fang, Changshuo Wang, Xiaoye Qu, Daizong Liu
机构
*
School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件学院)
;
Nanyang Technological University, Singapore(新加坡南洋理工大学)
;
University College London(伦敦大学学院)
;
Huazhong University of Science and Technology(华中科技大学)
;
Wuhan University(武汉大学)
机构
*
School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)
;
Nanyang Technological University, Singapore(新加坡南洋理工大学)
;
University College London(伦敦大学学院)
;
Guangzhou University(广州大学)
;
Wuhan University(武汉大学)
;
Nanjing University(南京大学)
OralAgent: Integrating Reasoning, Tools, and Knowledge for Interactive Dental Image Analysis
OralAgent: 融合推理、工具与知识的交互式牙科影像分析
Jing Hao, Siyuan Dai, Yongxin Zhang, Yuci Liang, Jiamin Wu, Jiahao Bao, Yuxuan Fan, Zanting Ye, Yanpeng Sun, Xinyu Zhang, Ming Hu, Liang Zhan, James Kit Hon Tsoi, Linlin Shen, Junjun He, Kuo Feng Hung
机构
*
Faculty of Dentistry, the University of Hongkong, Hong Kong SAR, China(香港大学牙科学院,中国香港特别行政区)
;
Department of Electrical and Computer Engineering, University of Pittsburgh, Pittsburgh, PA, USA(匹兹堡大学电气与计算机工程系,美国宾夕法尼亚州匹兹堡)
;
Shenzhen University, China(深圳大学,中国)
;
Department of Craniomaxillofacial Surgery, Shanghai Ninth People’s Hospital, China(上海第九人民医院口腔颌面外科部,中国)
;
Nanyang technological University, Singapore(南洋理工大学,新加坡)
;
School of Biomedical Engineering, Southern Medical University, China(南方医科大学生物医学工程学院,中国)
;
Singapore University of Technology and Design, Singapore(新加坡科技设计大学,新加坡)
;
University of Auckland, new zealand(奥克兰大学,新西兰)
;
Shanghai Artificial Intelligence Laboratory , China(上海人工智能实验室,中国)
机构
*
Nanyang Technological University(南洋理工大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Illinois Chicago(伊利诺伊大学香槟分校)
;
Tsinghua University(清华大学)
;
Sun Yat-sen University(中山大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
XTransfer: Modality-Agnostic Few-Shot Model Transfer for Human Sensing at the Edge
XTransfer: 面向边缘人体感知的模态无关小样本模型迁移
Yu Zhang, Xi Zhang, Hualin Zhou, Xinyuan Chen, Shang Gao, Hong Jia, Jianfei Yang, Yuankai Qi, Tao Gu
机构
*
Macquarie University, Sydney, NSW, Australia(麦考瑞大学,悉尼,新南威尔士州,澳大利亚)
;
Nanyang Technological University, Singapore(南洋理工大学,新加坡)
;
The University of Auckland, Auckland, New Zealand(奥克兰大学,奥克兰,新西兰)
SaFeR-Steer: Evolving Multi-Turn MLLMs via Synthetic Bootstrapping and Feedback Dynamics
SaFeR-Steer:通过合成引导和反馈动力学进化多轮多模态大语言模型
Haolong Hu, Hanyu Li, Tiancheng He, Huahui Yi, An Zhang, Qiankun Li, Kun Wang, Yang Liu, Zhigang Zeng
机构
*
Huazhong University of Science and Technology(华中科技大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
West China Biomedical Big Data Center, Sichuan University(四川大学西部生物医学大数据中心)
;
School of Public Policy and Administration, Chongqing University(重庆大学公共政策与管理学院)
;
Nanyang Technological University(南洋理工大学)
AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models
AlphaForgeBench:用大型语言模型对端到端交易策略设计进行基准测试
Wentao Zhang, Mingxuan Zhao, Jincheng Gao, Jieshun You, Huaiyu Jia, Yilei Zhao, Bo An, Shuo Sun
机构
*
Nanyang Technological University(南洋理工大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Hong Kong Polytechnic University(香港理工大学)