Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediators
使用具有特质-反应中介的虚拟受访者进行心理测量项目验证
Sungjib Lim, Woojung Song, Eun-Ju Lee, Yohan Jo
机构
*
Graduate School of Data Science, Seoul National University(首尔国立大学数据科学研究生院)
;
Department of Communication, Seoul National University(首尔国立大学通信系)
;
Interdisciplinary Program in Artificial Intelligence, Seoul National University(首尔国立大学人工智能跨学科项目)
专题命中
评测与基准
:LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
Kalyani Government Engineering College(卡利尼政府工程学院)
;
IIIT Guwahati(古瓦哈提理工学院)
;
IIIT Delhi(德里理工学院)
;
BITS Pilani Hyderabad Campus(比什帕利 Hyderabad 分校)
;
University of South Carolina(南卡罗来纳大学)
;
NIT Silchar(西里 char 工程学院)
;
San José State University(桑乔斯州立大学)
;
UCLA(加州大学洛杉矶分校)
;
Washington State University(华盛顿州立大学)
;
Vishwakarma Institute of Information Technology(维斯瓦卡arma 信息科技学院)
;
Gandhi Institute for Technological Advancement(甘地技术进步研究所)
;
BITS Pilani Goa(比什帕利 Goa 分校)
;
Meta AI
;
Amazon AI(亚马逊AI)
专题命中
评测与基准
:LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL
AstroMind: A High-Fidelity Benchmark for Spacecraft Behavior Reasoning Based on Large Language Models
AstroMind:基于大型语言模型的航天器行为推理高保真基准
Hao Liu, Siyuan Yang, Qinglei Hu, Dongyu Li
机构
*
Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院)
;
KTH Royal Institute of Technology(皇家理工学院)
;
School of Automation Science and Electrical Engineering, Beihang University(北京航空航天大学自动化科学与电气工程学院)
;
School of Cyber Science and Technology, Beihang University(北京航空航天大学网络空间安全学院)
专题命中
评测与基准
:large language model(title);language model(title);分类 cs.CL
ORACAL: A Robust and Explainable Multimodal Framework for Smart Contract Vulnerability Detection with Causal Graph Enrichment
ORACAL: 一种基于因果图增强的鲁棒且可解释的智能合约漏洞检测多模态框架
Tran Duong Minh Dai, Triet Huynh Minh Le, M. Ali Babar, Van-Hau Pham, Phan The Duy
机构
*
Information Security Lab, University of Information Technology(信息安全部,信息科技大学)
;
Vietnam National University(越南国家大学)
;
School of Computer Science and Information Technology, Adelaide University(计算机科学与信息技术学院,阿德莱德大学)
专题命中
评测与基准
:LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG
Teaching large language models to reason like expert diagnosticians
教会大型语言模型像专家诊断医生一样推理
Thomas A. Buckley, Riccardo Conci, Peter G. Brodeur, Jason Gusdorf, Sourik Beltrán, Bita Behrouzi, Byron Crowe, Jacob Dockterman, Muzzammil Muhammad, Sarah Ohnigian, Andrew Sanchez, James A. Diao, Aashna P. Shah, Daniel Restrepo, Eric S. Rosenberg, Andrew S. Lea, Emily Glanton, Kimberly LeBlanc, Undiagnosed Diseases Network, Marinka Zitnik, Scott H. Podolsky, Zahir Kanjee, Raja-Elie E. Abdulnour, Jacob M. Koshy, Adam Rodman, Arjun K. Manrai
机构
*
Department of Biomedical Informatics, Harvard Medical School(哈佛医学院生物医学信息学系)
;
Department of Medicine, Beth Israel Deaconess Medical Center(贝塞斯达医院内科部)
;
The Mongan Institute, Massachusetts General Hospital(麻省总医院蒙根研究所)
;
Division of Gastroenterology, Brigham and Women’s Hospital(布里洛妇女医院胃肠病科)
;
Department of Medicine, Brigham and Women’s Hospital(布里洛妇女医院内科部)
;
Department of Medicine, Massachusetts General Hospital(麻省总医院内科部)
;
Department of Pathology, Massachusetts General Hospital(麻省总医院病理学部)
;
Department of Health Humanities and Bioethics, University of Rochester School of Medicine and Dentistry(罗切斯特大学医学院和牙科学院健康人文与生物伦理学部)
;
Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard University(哈佛大学凯普纳人工智能研究所)
;
Center for the History of Medicine, Countway Library of Medicine, Harvard Medical School(哈佛医学院医学史中心,考特维图书馆)
;
Department of Global Health and Social Medicine, Harvard Medical School(哈佛医学院全球健康与社会医学部)
;
Division of Pulmonary and Critical Care Medicine, Brigham and Women’s Hospital(布里洛妇女医院呼吸科和重症医学科)
专题命中
评测与基准
:large language model(title);language model(title);分类 cs.AI
AI总结
提出 Dr. CaBot 代理 AI 系统,通过生成基于初始病例描述的幻灯片演示来模拟专家诊断推理,并在 NEJM CPC 和 NIH 未诊断疾病网络病例上取得优于前沿模型的表现,同时发布 CPC-Bench 基准以促进临床 AI 发展。
ViroBench: Benchmarking Nucleotide Foundation Models on Viral Genomics Tasks
ViroBench:病毒基因组学任务中的核苷酸基础模型基准测试
Dongxin Ye, Fang Hu, Han Hu, Shu Hu, Yang Tan, Wanli Ouyang, Stan Z. Li, Jie Cui, Nanqing Dong
机构
*
Shanghai Innovation Institute Shanghai China(深圳河套学院)
;
University of Electronic Science
;
Fudan University Shanghai China
;
Shanghai Artificial Intelligence Laboratory Shanghai China
;
Institute of Infection
;
Health Fudan University Shanghai China
;
Shanghai Sci-Tech Inno Center for Infection \& Immunity Shanghai China
;
Shanghai Jiao Tong University Shanghai China
;
Shenzhen Loop Area Institute Shenzhen China
;
Chinese University of Hong Kong Hong Kong China
;
Westlake University Hangzhou China
;
Shanghai Innovation Institute
;
Fudan University
;
Shanghai Artificial Intelligence Laboratory
;
Shanghai Sci-Tech Inno Center for Infection \& Immunity
;
Shanghai Jiao Tong University
;
Shenzhen Loop Area Institute
;
Chinese University of Hong Kong
;
Westlake University
Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation
在你说话之前了解你:多轮对话中用于LLM个性化的用户状态建模
Jiani Luo, Xiaoyan Zhao, Yang Zhang, Shuyi Miao, Bingbing Xu, Stefan Konigorski, Tat-Seng Chua
机构
*
School of Computing, National University of Singapore(新加坡国立大学计算机学院)
;
School of Artificial Intelligence, Beihang University(北航人工智能学院)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
German Institute of Human Nutrition Potsdam-Rehbruecke(德国人类营养研究所波茨坦-雷赫布鲁克)
Comments11 pages. Single-author preprint. Supplementary case-study report (Graph Isomorphism algorithm proposal with three theorems, five conjectures, complete complexity analysis, and hard-instance evaluation) available at https://spockstein.github.io/prima/case-study-graph-isomorphism.html
机构
*
The Ohio State University(俄亥俄州立大学)
;
The University of Chicago(芝加哥大学)
;
University College London(伦敦大学学院)
;
University of Michigan(密歇根大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Case Western Reserve University(凯斯西储大学)
;
Amazon(亚马逊)
专题命中
评测与基准
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
即时系统的时代已到来:挑战与机遇
Shu Liu, Alexander Krentsel, Shubham Agarwal, Mert Cemri, Ziming Mao, Soujanya Ponnapalli, Alexandros G. Dimakis, Sylvia Ratnasamy, Matei Zaharia, Aditya Parameswaran, Ion Stoica
Do Image-Text Metrics Respect Semantic Invariances?
图像-文本度量是否尊重语义不变性?
Amit Agarwal, Hitesh Laxmichand Patel, Meizhu Liu, Jyotika Singh, Karan Dua, Hansa Meghwani, Matthew Rowe, Michael Avendi, Yassi Abbasi, Tao Sheng, Sujith Ravi, Dan Roth
机构
*
Faculty of Engineering, Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学工程学院、计算机科学与工程系)
;
Artificial Intelligence Innovation and Incubation Institute, Fudan University(复旦大学人工智能创新与孵化院)
;
Shanghai Academy of AI for Science(上海人工智能科学研究院)
专题命中
评测与基准
:LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
TriVAL: A Tri-Validation Framework for Faithful Automatic Optimization Modeling
TriVAL: 一种用于忠实自动优化建模的三重验证框架
Ziyang Fang, JinXi Wang, Jinghui Zhong, Yew-Soon Ong
机构
*
School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院)
;
Centre for Frontier AI Research, Agency for Science, Technology and Research(科技研究局前沿人工智能研究中心)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
专题命中
评测与基准
:LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI