Modeling Community Attitude through Reaction Tone: A Human-AI Collaborative Framework for Evaluating LLM Alignment with Linguistic Behaviors in Online Communities
通过反应语气建模社区态度:评估LLM与在线社区语言行为对齐的人机协作框架
Nuan Wen, Xuezhe Ma
机构
*
Information Sciences Institute University of Southern California(南加州大学信息科学研究所)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
Department of Medicinal Chemistry, Faculty of Pharmacy, Tehran University of Medical Sciences(药学系,泰赫兰医科大学)
;
Department of Computer Sciences, Faculty of Mathematics and Computer Sciences, Amir Kabir University of Technology(计算机科学系,阿米尔·卡比尔技术大学)
;
Department of Mathematical Sciences, Sharif University of Technology(数学科学系,沙菲克技术大学)
;
Department of Computer Sciences, Missouri University of Science and Technology(计算机科学系,密苏里科学与技术大学)
;
Department of Computer Engineering, Sharif University of Technology(计算机工程系,沙菲克技术大学)
;
Department of Faculty of Interdisciplinary Science and Technology, Tarbiat Modares University(跨学科科学与技术学院,塔里亚特莫达res大学)
;
Electronics Research Institute, Sharif University of Technology(电子研究所,沙菲克技术大学)
;
Department of Electrical Engineering, Sharif University of Technology(电气工程系,沙菲克技术大学)
;
The Alan Turing Institute, London, United Kingdom(艾伦·图灵研究所,伦敦,英国)
;
Department of Radiation Oncology, Massachusetts General Hospital & Harvard Medical School(放射肿瘤科,麻省总医院及哈佛医学院)
;
Health Informatics Lab, Metropolitan College, Boston University(健康信息学实验室,波士顿大学)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Comments14 pages, 2 figures, 2 tables. The revised version includes McNemar's paired statistical analysis, Wilson confidence intervals, expanded methodological clarifications, a revised discussion of evidence retrieval, improved reproducibility details, and updated limitations
A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework
对LLM-as-a-Judge在医疗领域的综述及MedJUDGE框架
Chenyu Li, Zohaib Akhtar, Mingu Kwak, Yuelyu Ji, Hang Zhang, Tracey Obi, Yufan Ren, Xizhi Wu, Sonish Sivarajkumar, Harold P. Lehmann, Shyam Visweswaran, Michael J. Becich, Danielle L. Mowery, Renxuan Liu, Haoyang Sun, Yanshan Wang
机构
*
Department of Biomedical Informatics, School of Medicine, University of Pittsburgh(匹兹堡大学医学院生物医学信息学系)
;
Department of Health Information Management, School of Health and Rehabilitation Sciences, University of Pittsburgh(匹兹堡大学健康与康复科学学院健康信息管理系)
;
OpenCura, Health Innovation Consortium(OpenCura健康创新联盟)
;
Northwestern University, Kellogg School of Management(西北大学凯洛格管理学院)
;
Intelligent Systems Program, School of Computing and Information, University of Pittsburgh(匹兹堡大学计算与信息学院智能系统项目)
;
Johns Hopkins University School of Medicine Biomedical Informatics and Data Science(约翰霍普金斯大学医学院生物医学信息学与数据科学)
;
Clinical and Translational Science Institute, University of Pittsburgh(匹兹堡大学临床与转化科学研究所)
;
Institute for Biomedical Informatics, University of Pennsylvania(宾夕法尼亚大学生物医学信息学研究所)
;
Data Science, School of Computing and Information, University of Pittsburgh(匹兹堡大学计算与信息学院数据科学)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Learning to Ask: When LLM Agents Meet Unclear Instruction
学习提问:当LLM代理遇见模糊指令
Wenxuan Wang, Juluan Shi, Zixuan Ling, Yuk-Kit Chan, Chaozheng Wang, Cheryl Lee, Youliang Yuan, Jen-tse Huang, Wenxiang Jiao, Michael R. Lyu
机构
*
Renmin University of China(中国人民大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学深圳校区)
;
Johns Hopkins University(约翰霍普金斯大学)
;
Xiaohongshu Inc.(小红书公司)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Tell Me Why: Designing an Explainable LLM-based Dialogue System for Student Problem Behavior Diagnosis
告诉我原因:设计一个可解释的基于LLM的对话系统用于学生问题行为诊断
Zhilin Fan, Deliang Wang, Penghe Chen, Yu Lu
机构
*
School of Educational Technology, Beijing Normal University, Beijing, China(北京师范大学教育技术学院)
;
Advanced Innovation Center for Future Education, Beijing Normal University, Beijing, China(北京师范大学未来教育创新中心)
;
Faculty of Education, The University of Hong Kong, Hong Kong SAR, China(香港大学教育学院)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Jingyu Peng, Maolin Wang, Nan Wang, Jiatong Li, Yuchen Li, Yuyang Ye, Wanyu Wang, Pengyue Jia, Kai Zhang, Xiangyu Zhao
机构
*
City University of Hong Kong(香港城市大学)
;
Rutgers University(罗格斯大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Universiteit van Amsterdam(阿姆斯特丹大学)
;
Baidu Inc(百度公司)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
LLM Agents Predict Social Media Reactions but Do Not Outperform Text Classifiers: Benchmarking Simulation Accuracy Using 120K+ Personas of 1511 Humans
LLM代理预测社交媒体反应但不优于文本分类器:使用120,000多个1511人的人设进行基准测试
Ljubisa Bojic, Alexander Felfernig, Bojana Dinic, Velibor Ilic, Achim Rettinger, Vera Mevorah, Damian Trilling
机构
*
Institute for Artificial Intelligence Research and Development of Serbia(塞尔维亚人工智能研究与开发研究所)
;
University of Belgrade, Institute for Philosophy and Social Theory, Digital Society Lab(贝尔格莱德大学,哲学与社会理论研究所,数字社会实验室)
;
Complexity Science Hub, Vienna, Austria(维也纳复杂科学中心)
;
Graz University of Technology, Graz, Austria(技术大学格拉茨,格拉茨,奥地利)
;
University of Novi Sad, Faculty of Philosophy, Novi Sad, Serbia(诺维萨德大学,哲学学院,诺维萨德,塞尔维亚)
;
University of Trier, Trier, Germany(特里尔大学,特里尔,德国)
;
Vrije University Amsterdam, Amsterdam, Netherlands(阿姆斯特丹自由大学,阿姆斯特丹,荷兰)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Comments8 pages for main paper (exclude citation pages), 6 pages for appendix, totally 10 figures 7 tables and 2 algorithms. The paper is accepted by WACV 2026
Journal refIEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026
TriP-LLM: A Tri-Branch Patch-wise Large Language Model Framework for Time-Series Anomaly Detection
Yuan-Cheng Yu, Yen-Chieh Ouyang, Chun-An Lin
专题命中
评测与基准
:LLM(title,abstract);large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG
CommentsAccepted version of the paper published in IEEE Access (2025). Licensed under a Creative Commons Attribution 4.0 License (CC BY 4.0). Published version available at IEEE Xplore
Journal refIEEE Access, vol. 13, pp. 168643-168653, 2025