From Benchmarks to Skills: Low-Rank Factors for LLM Evaluation
从基准到技能:LLM评估的低秩因子
Aviya Maimon, Amir DN Cohen, Gal Vishne, Shauli Ravfogel, Reut Tsarfaty
机构
*
Bar-Ilan University(巴伊兰大学)
;
OriginAI
;
Data Science Institute Columbia University(哥伦比亚大学数据科学学院)
;
Center for Data Science New York University(纽约大学数据科学中心)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL
LLM-Powered Personalized Glycemic Assessment in Type 2 Diabetes with Wearable Sensor Data
基于可穿戴传感器数据的2型糖尿病个性化血糖评估:LLM驱动方法
Yifan Gao, Yanmin Gong, Yun Shi, Yuanxiong Guo
机构
*
Department of Information Systems and Cybersecurity, The University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校信息系统与网络安全系)
;
School of Engineering Medicine, Texas A&M University(德克萨斯农工大学工程医学院)
;
Department of Family and Community Medicine, The University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校家庭与社区医学系)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
CommentsAccepted at the 2nd Workshop on Misinformation Detection in the Era of LLMs (MisD), co-located with ICWSM 2026, May 26, 2026, Los Angeles, CA, USA
Journal refProceedings of the ICWSM Workshops, MisD 2026: The 2nd Workshop on Misinformation Detection in the Era of LLMs, 2026
Constrained Semantic Decompression in LLMs through Persian Proverb-Conditioned Story Generation
通过波斯谚语条件故事生成实现LLM中的约束语义解压缩
Zahra Habibzadeh, Paria Khoshtab, Amir Mesbah, Yadollah Yaghoobzadeh
机构
*
Tehran Institute for Advanced Studies, Khatam University, Iran(德黑兰高级研究所,卡塔姆大学,伊朗)
;
School of Electrical and Computer Engineering, College of Engineering, University of Tehran, Tehran, Iran(电气与计算机工程学院,工程学院,德黑兰大学,伊朗)
专题命中
评测与基准
:LLM(title_cn,abstract);large language model(abstract);language model(abstract);prompting(abstract)
RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue
RogueAI: 一种用于检测对话中授权AI欺骗的逆向图灵测试
Sara Candussio, Emanuele Ballarin, Lorenzo Bonin, Sandro Junior Della Rovere, Luca Bortolussi
机构
*
AILab, MIGe, University of Trieste(的里雅斯特大学)
;
Computational Statistics and Machine Learning, Istituto Italiano di Tecnologia(意大利理工学院)
;
DIA, University of Trieste(的里雅斯特大学)
专题命中
评测与基准
:LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL
Andy Wang, Parv Mahajan, David Demitri Africa, Alexandra Souly, Jordan Taylor, Robert Kirk
机构
*
Constellation University of Wisconsin-Madison(威斯康星大学麦迪逊分校星座研究所)
;
Constellation Georgia Institute of Technology(佐治亚理工学院星座研究所)
;
UK AI Security Institute(英国人工智能安全研究所)
专题命中
评测与基准
:language model(title,abstract);large language model(title);分类 cs.AI
CommentsSubmitted to arXiv. 20 pages, 4 figures. Work on emotion-based prompt engineering for text-to-image diffusion models with applications in personalized image generation
机构
*
Tianqiao and Chrissy Chen Institute(天桥和克里斯西·陈研究所)
;
EverMind AI Inc.(EverMind AI公司)
;
Shanghai Mental Health Center, Shanghai Jiao Tong University School of Medicine(上海精神卫生中心,上海交通大学医学院)
机构
*
School of Artificial Intelligence, Shanghai Jiao Tong University, Shanghai, China(上海交通大学人工智能学院)
;
Zhongguancun Academy, Beijing, China(中关村学院)
;
School of Integrated Circuits, Shanghai Jiao Tong University, Shanghai, China(上海交通大学集成电路学院)
;
School of Computer Science, Shanghai Jiao Tong University, Shanghai, China(上海交通大学计算机科学学院)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University, Beijing, China(北京大学计算机科学学院多媒体信息处理国家重点实验室)