Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
机构 * ECE Department(电子工程系) ; Carnegie Mellon University(卡内基梅隆大学) ; Computer Science Department(计算机科学系) ; Columbia University(哥伦比亚大学) ; Rotman School of Management(罗特曼管理学院) ; University of Toronto(多伦多大学) ; Department of Statistics(统计学系) ; George Washington University(乔治华盛顿大学) ; Department of Computer Science(计算机科学系) ; San Francisco State University(旧金山州立大学)
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG