"I Don't Know" -- Towards Appropriate Trust with Certainty-Aware Retrieval Augmented Generation
我不知道--迈向具有确定性意识的适当信任
Daan Di Scala, Maaike de Boer, Pınar Yolum
机构
*
TNO Netherlands Organisation for Applied Scientific Research, Department Data Science(荷兰应用科学研究院,数据科学部门)
;
Utrecht University, Department of Information and Computing Sciences(乌得勒支大学,信息与计算科学系)
专题命中
评测与基准
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI
GDPR Auto-Formalization with AI Agents and Human Verification
GDPR自动形式化与AI代理和人类验证
Ha Thanh Nguyen, Wachara Fungwacharakorn, Sabine Wehnert, May Myo Zin, Yuntao Kong, Jieying Xue, Michał Araszkiewicz, Randy Goebel, Ken Satoh
机构
*
Center for Juris-Informatics, ROIS-DS(法律信息中心,ROIS-DS)
;
Ruhr-University Bochum, RC-Trust(波恩鲁尔大学,RC-Trust)
;
Alberta Machine Intelligence Institute, University of Alberta(阿尔伯塔机器智能研究所,阿尔伯塔大学)
专题命中
评测与基准
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI
TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation
TRIP-Evaluate: 一个用于评估交通领域大模型的开放多模态基准
Han Gong, Zhen Zhou, Yunyang Shi, Yan Tan, Jinbiao Huo, Qi Hong, Zhiyuan Liu
机构
*
School of Transportation(交通学院)
;
Southeast University(东南大学)
;
School of Artificial Intelligence and Computer Science(人工智能与计算机科学学院)
;
Jiangnan University(江南大学)
;
Department of Civil and Environmental Engineering(土木与环境工程系)
;
Hong Kong Polytechnic University(香港理工大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract);prompting(abstract);分类 cs.AI、cs.LG
Validation of Whole-Slide Foundation Models for Image Retrieval in TCGA Data
对TCGA数据中整张滑动图像检索的整张滑动图像基础模型进行验证
Tianhao Lei, Parsa Esmaeilkhani, Saghir Alfasly, Wataru Uegami, Judy C. Boughey, Matthew P. Goetz, Krishna R. Kalari, H. R. Tizhoosh
机构
*
KIMIA Lab, Artificial Intelligence and Informatics, Mayo Clinic, Rochester, MN, USA(KIMIA实验室,人工智能与信息学,梅奥诊所,罗切斯特,明尼苏达州,美国)
;
Department of Neurology, Northwestern University Feinberg School of Medicine, Chicago, IL, USA(神经病学系,北western大学费因伯格医学院,芝加哥,伊利诺伊州,美国)
;
Department of Computer Science, Temple University, Philadelphia, PA, USA(计算机科学系,泰勒大学,费城,宾夕法尼亚州,美国)
;
Department of Breast and Melanoma Surgical Oncology, Comprehensive Cancer Center, Mayo Clinic, Rochester, MN, USA(乳腺和黑色素瘤外科肿瘤学系,综合癌症中心,梅奥诊所,罗切斯特,明尼苏达州,美国)
;
Department of Oncology, Comprehensive Cancer Center, Mayo Clinic, Rochester, MN, USA(肿瘤学系,综合癌症中心,梅奥诊所,罗切斯特,明尼苏达州,美国)
;
Department of Quantitative Health Sciences, Mayo Clinic, Rochester, MN, USA(定量健康科学系,梅奥诊所,罗切斯特,明尼苏达州,美国)
Breaking the Silence: A Dataset and Benchmark for Bangla Text-to-Gloss Translation
打破沉默:Bangla文本到词组翻译的数据集和基准
Sharif Mohammad Abdullah, Abhijit Paul, Shubhashis Roy Dipta, Zarif Masud, Shebuti Rayana, Ahmedul Kabir
机构
*
IIT, University of Dhaka, Bangladesh(达卡大学理工学院,孟加拉国)
;
University of Maryland, Baltimore County, USA(马里兰大学巴尔的摩县分校,美国)
;
SUNY, Old Westbury, USA(SUNY,美国奥尔德韦斯特伯里)
AMSnet-q: Unsupervised Circuit Identification and Performance Labeling for AMS Circuits
AMSnet-q:无监督的电路识别与性能标注用于AMS电路
Ze Zhang, Junzhuo Zhou, Yichen Shi, Zhuofu Tao, Rui Ji, Zhiping Yu, Quan Chen, Ting-Jung Lin, Lei He
机构
*
Southern University of Science and Technology(南方科技大学)
;
University of California Los Angeles(加州大学洛杉矶分校)
;
Tsinghua University(清华大学)
;
Eastern Institute of Technology Ningbo(宁波东部技术研究院)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI
机构
*
School of IC, Peking University, Beijing, China(北京大学信息科学技术学院)
;
College of Artificial Intelligence, Xi'an Jiaotong University, Xi'an, China(西安交通大学人工智能学院)
;
Department of Precision Instruments, Tsinghua University, Beijing, China(清华大学精密仪器系)
;
School of Software and Microelectronics, Peking University, Beijing, China(北京大学软件与微电子学院)
;
School of Computer Science, Peking University, Beijing, China(北京大学计算机科学系)
;
Institute of EDA, Peking University, Beijing, China(北京大学EDA研究院)
;
Beijing Advanced Innovation Center for IC, Beijing, China(北京集成电路先进创新中心)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI
A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis
多视角媒体画像套件:资源、评估与分析
Muhammad Arslan Manzoor, Dilshod Azizov, Daniil Orel, Umer Siddique, Zain Muhammad Mujahid, Yufang Hou, Preslav Nakov
机构
*
MBZUAI(穆巴扎人工智能研究院)
;
Interdisciplinary Transformation University(跨学科转型大学)
;
University of Texas at San Antonio, USA(德克萨斯大学圣安东尼奥分校)
;
University of Copenhagen, Denmark(哥本哈根大学)
机构
*
Department of Computer Science, University of British Columbia, Vancouver, Canada(英属哥伦比亚大学计算机科学系)
;
Department of Artificial Intelligence, IIT Hyderabad, Hyderabad, India(印度海得拉巴理工学院人工智能系)
;
Amazon, Delhi, India(印度德里亚马逊公司)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI
机构
*
School of Interactive Computing, Georgia Institute of Technology(佐治亚理工学院交互计算学院)
;
College of Computing, Georgia Institute of Technology(佐治亚理工学院计算机学院)
;
School of Computing Technologies, RMIT University(皇家墨尔本理工大学计算技术学院)
机构
*
Sangfor Technologies Inc.(Sangfor 技术公司)
;
Wuhan University(武汉大学)
;
The University of Melbourne(墨尔本大学)
;
Dalian Maritime University(大连海事大学)
;
Hunan University of Technology and Business(湖南工业大学;湘江实验室)
;
Xiangjiang Laboratory
Ali Ezzat Shahroor, Mohamed Bayan Kmainasi, Abul Hasnat, Dimitar Dimitrov, Giovanni Da San Martino, Preslav Nakov, Firoj Alam
机构
*
Qatar Computing Research Institute(卡塔尔计算研究所)
;
Sofia University "St. Kliment Ohridski"(索菲亚大学"圣克莱孟·奥赫里迪斯")
;
University of Padova(帕多瓦大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
Geospatial foundation-model embeddings improve population estimation unevenly across space and scale
地理基础模型嵌入物在空间和尺度上不均等地提高人口估计
Wenbin Zhang, Eimear Cleary, Francisco Rowe, Somnath Chaudhuri, Maksym Bondarenko, Shengjie Lai, Andrew J. Tatem
机构
*
WorldPop, School of Geography and Environmental Sciences, University of Southampton, United Kingdom(世界人口研究机构,地理与环境科学学院,南安普顿大学,英国)
;
Geographic Data Science Lab, Department of Geography and Planning, School of Environmental Sciences, University of Liverpool, United Kingdom(地理数据科学实验室,地理与规划系,环境科学学院,利物浦大学,英国)
Retrieval-Guided Generation for Safer Histopathology Image Captioning
基于检索的生成用于更安全的病理科图像描述生成
Md. Enamul Hoq, Wataru Uegami, Saghir Alfasly, Ghazal Alabtah, Sahar Rahimi Malakshan, Armita Kazemi, Alex T. Schmitgen, Fred Prior, H. R. Tizhoosh
机构
*
Kimia Lab, Department of Artificial Intelligence \& Informatics, Mayo Clinic, Rochester, MN, USA
;
Department of Biomedical Informatics, University of Arkansas for Medical Sciences, Little Rock, AR, USA
;
Lane Department of Computer Science
;
Electrical Engineering, West Virginia University, Morgantown, WV, USA
;
Department of Computer Science
;
Engineering, Princeton University, Princeton, NJ, USA
;
Department of Computer Sciences, University of Wisconsin--Madison, Madison, WI, USA
机构
*
Department of Computer Science, University of North Carolina at Charlotte(北卡罗来纳大学夏洛特分校计算机科学系)
;
Department of Computer Science, Dartmouth College(达特茅斯学院计算机科学系)
;
School of Computing, University of Georgia(佐治亚大学计算科学学院)
;
School of Electrical and Computer Engineering, University of Georgia(佐治亚大学电气与计算机工程学院)