Knowledge Base Construction for Knowledge-Augmented Text-to-SQL
机构 * KAIST(韩国科学技术院) ; IBM Research(IBM研究院)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments ACL Findings 2025
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * KAIST(韩国科学技术院) ; IBM Research(IBM研究院)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments ACL Findings 2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI
Comments Accepted by ICML 2025
机构 * Carnegie Mellon University(卡内基梅隆大学) ; MBZUAI(穆斯林人工智能研究院)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
机构 * Spiral Works ; Univ. Illinois, Urbana-Champaign(伊利诺伊大学,厄巴纳-香槟分校) ; University of Michigan(密歇根大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments Accepted at ICCC 2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG
Comments We are currently revising the methodology described in the manuscript to improve its clarity. We have decided to withdraw the current version until a more robust and complete version is ready
机构 * Department of Information Engineering, University of Padova(信息工程系,帕多瓦大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.LG
Comments Accepeted as long paper at "The 3rd Workshop for Out-of-Distribution Generalization in Computer Vision Foundation Models", ECCV 2024
Journal ref ECCV 2024 Workshops. ECCV 2024. Lecture Notes in Computer Science
机构 * RMIT University(拉筹伯大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG
Comments 11 pages, 8 figures, submitted to IEEE Transactions on Visualization and Computer Graphics (TVCG)
机构 * Tencent YouTu Lab(腾讯优图实验室) ; Siemens AG(西门子股份公司) ; Technical University of Munich(慕尼黑技术大学) ; Shanghai Jiao Tong University(上海交通大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
Comments 27 pages, 15 figures, 22 tables
机构 * Department of Electrical and Computer Engineering(电气与计算机工程系) ; University of Rochester(罗切斯特大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
Comments 8 pages, 5 figures, not submitted to any conference
机构 * Tencent YouTu Lab(腾讯优图实验室)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
Comments Accepted by NeurIPS 2024
机构 * International Computer Science Institute(国际计算机科学研究所) ; School of Computing and Augmented Intelligence(计算与增强智能学院) ; Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室) ; Department of Statistics(统计学系) ; University of California at Berkeley(加州大学伯克利分校)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
机构 * ProjectEndgame
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments 20 pages, 8 figures, 18 tables
机构 * Computer Vision Laboratory, ETH Zürich(苏黎世联邦理工学院计算机视觉实验室) ; Integrated System Laboratory, ETH Zürich(苏黎世联邦理工学院集成系统实验室) ; School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院)
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.CV、cs.AI
Comments Technical report. Accepted by Visual Intelligence. Code is released at https://github.com/zhoustan/SAM2-VCOS
机构 * Department of Electronic and Computer Engineering, HKUST(香港科技大学电子与计算机工程系) ; Department of Radiology, Guangdong Provincial Key Laboratory of Malignant Tumor Epigenetics and Gene Regulation, Sun Yat-Sen Memorial Hospital, Sun Yat-Sen University(中山大学放射科、广东省恶性肿瘤表观遗传与基因调控重点实验室、中山纪念医院) ; Department of Computer Science and Engineering, HKUST(香港科技大学计算机科学与工程系)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments International Conference on Machine Learning
机构 * University of Minnesota(明尼苏达大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments Project page: https://minnesotanlp.github.io/scitalk-project-page/
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
Comments Project page: https://describe-anything.github.io/
机构 * FAIR at Meta(Meta 的 FAIR)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.LG
Comments Code and Data: https://github.com/HaroldChen19/VistaDPO
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments 55 pages
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments 19 pages
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
Comments 9 pages, 6 figures
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI
Comments CVPR 2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.CV、cs.LG
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
Comments IEEE/CVF Computer Vision and Pattern Recognition 2025; 22 pages
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI、cs.LG
Comments 2023 IEEE International Conference on Data Mining Workshops (ICDMW)
Journal ref 2023 IEEE International Conference on Data Mining Workshops (ICDMW)