From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning
从提示到路面通过时间:代理场景到计划推理中的时间定位
Ahmed Y. Gado, Omar Y. Goba, Alaa Hassanein, Catherine M. Elias, Ahmed Hussein
机构
*
Computer Science & Engineering Department, German University in Cairo (GUC), Egypt(德国亚历山大大学(GUC)计算机科学与工程系,埃及)
;
C-DRiVeS Lab: Cognitive Driving Research in Vehicular Systems, Cairo, Egypt(认知驾驶系统实验室,埃及开罗,C-DRiVeS)
;
M.Eng. Robotics Candidate at Deggendorf Institute of Technology, Germany(德国德格多夫技术学院机器人硕士候选人)
;
IAV GmbH, Berlin, Germany(德国柏林IAV GmbH公司)
Grounding Large Language Models as Generalizable Policies in Network Control
大语言模型作为网络优化的通用策略
Duo Wu, Linjia Kang, Zhimin Wang, Fangxin Wang, Wei Zhang, Chongbo Sun, Xuefeng Tao, Wei Yang, Le Zhang, Wenwu Zhu, Peng Cui, Zhi Wang
机构
*
Bytedance(字节跳动)
;
Shenzhen International Graduate School(深圳国际研究生院)
;
Tsinghua University(清华大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Department of Computer Science and Technology(计算机科学与技术系)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
Omni-Persona:系统性基准测试与改进多模态个性化
Yeongtak Oh, Dongwook Lee, Sangkwon Park, Heeseung Kim, Sungroh Yoon
机构
*
Department of Electrical and Computer Engineering, Seoul National University(首尔国立大学电气与计算机工程系)
;
Interdisciplinary Program in Artificial Intelligence, Seoul National University(首尔国立大学人工智能跨学科项目)
;
Department of Artificial Intelligence, University of Seoul(首尔大学人工智能系)
专题命中
视觉定位与Grounding
:grounding(abstract);multimodal large language model(abstract);分类 cs.CV
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
Eastern Institute of Technology, Ningbo(宁波东方理工大学)
;
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
;
Shanghai Jiao Tong University(上海交通大学)