Can Foundation Models Hear What Made That Sound? A Tiered Benchmark of Audio-Language Models and Traditional Classifiers for Closed-Set Sound Source Identification
基础模型能否识别声音的来源?针对闭集声源识别的音频-语言模型与传统分类器的分层基准测试
Sajjad Abdoli, Ghassan Al-Sumaidaee, Ahmad ElShiekh, Ahmed Rashad
机构
*
Department of Computer Science and Engineering (AI & ML), Kakatiya Institute of Technology and Science(卡凯蒂亚理工学院计算机科学与工程系(人工智能与机器学习))
;
Department of Computer Science and Engineering, Kakatiya Institute of Technology and Science(卡凯蒂亚理工学院计算机科学与工程系)
;
Boston University(波士顿大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI
机构
*
TJUNLP Lab, School of Computer Science and Technology, Tianjin University(天津大学计算机科学与技术学院TJUNLP实验室)
;
The International Joint Institute of Tianjin University(天津大学国际联合研究院)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.CL
机构
*
Universitas Universal(印尼环球大学)
;
Beijing Language and Culture University(北京语言大学)
;
Tianjin University(天津大学)
;
School of Artificial Intelligence, Tianjin University(天津大学人工智能学院)
;
School of Computer Science and Technology, Tianjin University(天津大学计算机科学与技术学院)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.CL
机构
*
School of Environment, Tsinghua University(清华大学环境学院)
;
College of Economics and Management, Beijing University of Technology(北京工业大学经济与管理学院)
;
State Key Laboratory of Iron and Steel Industry Environmental Protection, School of Environment, Tsinghua University(清华大学环境学院钢铁工业环境保护国家重点实验室)
;
Appraisal Center for Environmental Engineering, Ministry of Ecology and Environment(生态环境部环境工程评估中心)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI
EduArt: An educational-level benchmark for evaluating art history knowledge in large language models
EduArt:评估大型语言模型艺术史知识的教育级基准
Gianmarco Spinaci, Lukas Klic, Giovanni Colavizza
机构
*
University of Bologna(博洛尼亚大学)
;
Villa i Tatti – The Harvard University Center for Italian Renaissance Studies(哈佛大学意大利文艺复兴研究中心(I Tatti))
;
University of Copenhagen(哥本哈根大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.CL
A Systematic Evaluation of Black-Box Uncertainty Estimation Methods for Large Language Models
大型语言模型黑盒不确定性估计方法的系统评估
Jiayi Wang, Xu-Yao Zhang
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所多模态人工智能系统国家重点实验室)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI
UrbanWell: Benchmarking Multimodal Large Language Models for Spatio-Temporal Urban Wellbeing Analytics
UrbanWell: 面向时空城市福祉分析的多模态大语言模型基准测试
Yanxin Xi, Xiang Su, Jie Feng, Yu Liu, Sasu Tarkoma, Pan Hui
机构
*
University of Helsinki(赫尔辛基大学)
;
Zhongguancun Academy(中关村学院)
;
University of Oxford(牛津大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI
Constrained Semantic Decompression in LLMs through Persian Proverb-Conditioned Story Generation
通过波斯谚语条件故事生成实现LLM中的约束语义解压缩
Zahra Habibzadeh, Paria Khoshtab, Amir Mesbah, Yadollah Yaghoobzadeh
机构
*
Tehran Institute for Advanced Studies, Khatam University, Iran(德黑兰高级研究所,卡塔姆大学,伊朗)
;
School of Electrical and Computer Engineering, College of Engineering, University of Tehran, Tehran, Iran(电气与计算机工程学院,工程学院,德黑兰大学,伊朗)
专题命中
评测与基准
:LLM(title_cn,abstract);large language model(abstract);language model(abstract);prompting(abstract)
Evidence-Based Intelligent Diagnostic and Therapeutic Visualization System with Large Language Models: Multi-Turn Interaction and Multimodal Treatment Plan Generation
基于证据的智能诊断与治疗可视化系统与大语言模型:多轮交互与多模态治疗方案生成
Yunhan Wang, Yuda Wang, Zhiying Tu, Mingqiang Song, Li Song, Kun Li, Dianhui Chu, Bolin Zhang
机构
*
Harbin Institute of Technology, Weihai(哈尔滨工业大学(威海))
;
Harbin Institute of Technology (Weihai) Qingdao Research Institute(哈尔滨工业大学(威海)青岛研究院)
;
Shandong Key Laboratory of Digital Service Computing Technology and Systems(山东省数字服务计算技术与系统重点实验室)
;
Weihai Municipal Hospital(威海市人民医院)
;
Shanghai Taizhu Technology Co., Ltd(上海泰山技术有限公司)
;
Tianjin Zhifu Qihuang Medical Technology Co., Ltd(天津中孚启黄医疗技术有限公司)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI
Once a Response, Always a Response: Detecting LLM-generated Text via Latent Prompt Restoration
一次响应,始终是响应:通过潜在提示恢复检测大语言模型生成的文本
Hongrui Bao, Yubing Ren, Yanan Cao, Jinhan You, Fang Fang, Shi Wang
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
专题命中
评测与基准
:LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI