机构
*
Northwestern University in Qatar(卡塔尔西北大学)
;
Independent Researcher(独立研究员)
;
Hamad Bin Khalifa University(哈马德·本·卡伊夫大学)
;
University of Tübingen(图宾根大学)
专题命中
视觉定位与Grounding
:vision language model(abstract);grounding(abstract)
Early Risk Prediction with Temporally and Contextually Grounded Clinical Language Processing
利用时序和情境 grounding 的临床语言处理进行早期风险预测
Rochana Chaturvedi, Yue Zhou, Andrew D. Boyd, Brian T. Layden, Mudassir Rashid, Lu Cheng, Ali Cinar, Barbara Di Eugenio
机构
*
Kellogg School of Management, Northwestern University(西北大学凯洛格管理学院)
;
University of Illinois Chicago(伊利诺伊大学芝加哥分校)
;
Illinois Institute of Technology(伊利诺伊理工学院)
CLASP: Closed-loop Asynchronous Spatial Perception for Open-vocabulary Desktop Object Grasping
CLASP: 闭环异步空间感知用于开放词汇桌面物体抓取
Yiran Ling, Wenxuan Li, Siying Dong, Yize Zhang, Xiaoyao Huang, Jing Jiang, Ruonan Li, Jie Liu
机构
*
Harbin Institute of Technology(哈尔滨工业大学)
;
National Key Laboratory of Smart Farm Technologies and Systems(智慧农场技术与系统全国重点实验室)
;
Peng Cheng Laboratory(鹏城实验室)
;
Northeastern University(东北大学)
A Semantic Observer Layer for Autonomous Vehicles: Pre-Deployment Feasibility Study of VLMs for Low-Latency Anomaly Detection
面向自动驾驶的语义观察层:VLMs在低延迟异常检测中的预部署可行性研究
Kunal Runwal, Swaraj Gajare, Daniel Adejumo, Omkar Ankalkope, Siddhant Baroth, Aliasghar Arab
机构
*
The City College of New York, Grove School of Engineering(纽约市立学院格罗夫工程学院)
;
Department of Mechanical and Aerospace Engineering, Tandon School of Engineering, New York University(纽约大学坦登工程学院机械与航空航天工程系)
;
Department of Electrical and Computer Engineering, Tandon School of Engineering, New York University(纽约大学坦登工程学院电气与计算机工程系)
From Scale to Speed: Adaptive Test-Time Scaling for Image Editing
从尺度到速度:面向图像编辑的自适应测试时间缩放
Xiangyan Qu, Zhenlong Yuan, Jing Tang, Rui Chen, Datao Tang, Meng Yu, Lei Sun, Yancheng Bai, Xiangxiang Chu, Gaopeng Gou, Gang Xiong, Yujun Cai
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院)
;
AMAP, Alibaba Group(阿里巴巴集团高德地图)
;
University of Queensland(昆士兰大学)
机构
*
The State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences, China(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
Spatiotemporal AI, China(时空人工智能,中国)
;
Hangzhou International Innovation Institute, Beihang University, China(杭州国际创新研究院,北航,中国)
;
Georgia Institute of Technology, China(佐治亚理工学院,中国)
;
Key Laboratory of Computing Power Network and Information Security, Ministry of Education(计算功率网络与信息安全重点实验室,教育部;山东省计算机科学中心,齐鲁工业大学(山东省科学院),中国)
;
Shandong Computer Science Center, Qilu University of Technology (Shandong Academy of Sciences), China
专题命中
视觉定位与Grounding
:grounding(abstract);multimodal large language model(abstract)