AI Knows When It's Being Watched: Functional Strategic Action and Contextual Register Modulation in Large Language Models
AI 知道它在被观察:大型语言模型中的功能性战略行为与情境语境调节
Vinicius Covas, Jorge Alberto Hidalgo Toledo
机构
*
Center for Applied Communication Research (CICA)(应用沟通研究中心)
;
Human & NonHuman Communication Laboratory(人类与非人类沟通实验室)
;
Faculty of Communication(传播学院)
;
Universidad Anáhuac México(墨西哥安纳胡阿克大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.CL、cs.AI
AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models
AgenticEval: 向大型语言模型的代理和自演化安全评估迈进
Yixu Wang, Xin Wang, Yang Yao, Xinyuan Li, Xibang Yang, Yan Teng, Xingjun Ma, Yingchun Wang
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Fudan University(复旦大学)
;
The University of Hong Kong(香港大学)
;
East China Normal University(华东师范大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(summary_cn);分类 cs.AI
Yahan Li, Jifan Yao, John Bosco S. Bunyi, Adam C. Frank, Angel Hsing-Chi Hwang, Ruishan Liu
机构
*
Department of Computer Science, University of Southern California(南加州大学计算机科学系)
;
Department of Electrical and Computer Engineering, University of Southern California(南加州大学电气与计算机工程系)
;
Suzanne Dworak-Peck School of Social Work, University of Southern California(南加州大学苏兹安·德沃拉克-佩克社会工作学院)
;
Department of Psychiatry and the Behavioral Sciences, University of Southern California(南加州大学精神病学与行为科学系)
;
Annenberg School for Communication, University of Southern California(南加州大学安纳伯格通信学院)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.CL
RxEval: A Prescription-Level Benchmark for Evaluating LLM Medication Recommendation
RxEval: 一个处方级基准,用于评估LLM药物推荐
Shuhao Chen, Weisen Jiang, Changmiao Wang, Xiaoqing Wu, Xuanren Shi, Yu Zhang, James T. Kwok
机构
*
The Hong Kong University of Science and Technology(香港科技大学)
;
Southern University of Science and Technology(南方科技大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学深圳校区)
;
Shenzhen University General Hospital(深圳大学人民医院)
CommentsThe paper has mistake of undertaking political spaces to semantic dimensions. This needs to be removed because this is a fetal flaw in consideration. The initial hypothesis and premise needs to be rigorously formulated within the political landscape not generalizing the metrics. Hence a withdrawal for now is necessary
Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia
欺骗、检测与揭露:大型语言模型扮演迷你黑帮
Davi Bastos Costa, Renato Vicente
机构
*
TELUS Digital Research Hub(TELUS数字研究中心)
;
Center for Artificial Intelligence and Machine Learning(人工智能与机器学习中心)
;
Institute of Mathematics, Statistics and Computer Science(数学、统计与计算机科学研究所)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI