What Matters for Grocery Product Retrieval with Open Source Vision Language Models
在开源视觉语言模型中,什么因素影响杂货产品检索
Emmanuel G. Maminta, Rowel O. Atienza
机构
*
AI Graduate Program, University of the Philippines, Diliman, Quezon City(菲律宾大学达林学院人工智能研究生项目)
;
EEEI, University of the Philippines, Diliman, Quezon City(菲律宾大学达林学院电子工程系)
机构
*
University of Texas Austin(德克萨斯大学奥斯汀分校)
;
California Institute of Technology(加州理工学院)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Stanford University(斯坦福大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Microsoft Research(微软研究院)
;
Northwestern University(西北大学)
;
University of Cambridge(剑桥大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI
Conv-FinRe: A Conversational and Longitudinal Benchmark for Utility-Grounded Financial Recommendation
Conv-FinRe:一种用于实用导向财务推荐的对话和纵向基准
Yan Wang, Yi Han, Lingfei Qian, Yueru He, Xueqing Peng, Dongji Feng, Zhuohan Xie, Vincent Jim Zhang, Rosie Guo, Fengran Mo, Jimin Huang, Yankai Chen, Xue Liu, Jian-Yun Nie
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
Columbia University(哥伦比亚大学)
;
California State University(加州州立大学)
;
University of Montreal(蒙特利尔大学)
;
The University of Manchester(曼彻斯特大学)
;
McGill University(麦吉尔大学)
FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs
FinAuditing: 一个基于财务分类结构的多文档基准,用于评估LLMs
Yan Wang, Keyi Wang, Shanshan Yang, Jaisal Patel, Jeff Zhao, Fengran Mo, Xueqing Peng, Lingfei Qian, Yankai Chen, Víctor Gutiérrez-Basulto, Jimin Huang, Guojun Xiong, Xiao-Yang Liu, Xue Liu, Jian-Yun Nie
机构
*
Columbia University(哥伦比亚大学)
;
Stevens Institute of Technology(史蒂文斯理工学院)
;
Rensselaer Polytechnic Institute(拉特格斯理工学院)
;
University of Montreal(蒙特利尔大学)
;
McGill University(麦吉尔大学)
;
MBZUAI(麦吉尔大学人工智能研究所)
;
Cardiff University(卡迪夫大学)
;
The University of Manchester(曼彻斯特大学)
;
Harvard University(哈佛大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL
To Trust or Not to Trust: Authors' Response to AI-based Reviews
信任还是不信任:作者对基于AI的评论的回应
César Leblanc, Lukas Picek
机构
*
École Normale Supérieure(巴黎高等师范学院)
;
Sorbonne University(索邦大学)
;
University of West Bohemia(西波什埃大学)
;
Massachusetts Institute of Technology(麻省理工学院)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI
Comments19 pages, 7 figures, 8 tables, 50 references. A shortened workshop version has been submitted to the BPM 2026 Workshop. This preprint is the complete version
Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms
基于触觉的多模态融合在具身智能中的应用:视觉、语言和接触驱动范式的综述
Zhixiang Cao, Di Tian, Runwei Guan, Yanzhou Mu, Xiaolou Sun, Shaofeng Liang, Daizong Liu, Tao Huang, Yutao Yue, Henghui Ding, Bin Fang, Alex Zhou, Qing-Long Han, Hui Xiong
机构
*
School of Electronic Science and Engineering, Xi’an Jiaotong University, China(西安交通大学电子科学与技术学院)
;
Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou), China(香港科技大学(广州)人工智能研究所)
;
State Key Laboratory for Novel Software Technology, Nanjing University, China(南京大学新型软件技术国家重点实验室)
;
Purple Mountain Laboratory, China(紫金山实验室)
;
Institute for Math & AI, Wuhan University, China(武汉大学数学与人工智能学院)
;
Centre for AI and Data Science Innovation and the School of Science and Engineering, James Cook University, Australia(詹姆斯库克大学人工智能与数据科学创新中心及科学与工程学院)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications, China(北京邮电大学人工智能学院)
;
Institute of Big Data, Fudan University, China(复旦大学大数据研究院)
;
Linkerbot (Beijing) Technology Co., Ltd, China(北京链动科技有限公司)
;
School of Engineering, Swinburne University of Technology, Melbourne(斯威本技术大学工程学院)
专题命中
评测与基准
:large language model(abstract);language model(abstract)