In-context superposition: human-like working memory interference in large language models
类人工作记忆干扰在大语言模型中
Hua-Dong Xiong, Li Ji-An, Jiaqi Huang, Robert C. Wilson, Kwonjoon Lee, Xue-Xin Wei
机构
*
School of Psychological and Brain Sciences, Georgia Tech(佐治亚理工学院心理与脑科学学院)
;
Department of Psychology, New York University(纽约大学心理学系)
;
Department of Cognitive Science, Indiana University Bloomington(印第安纳大学布卢明顿分校认知科学系)
;
Honda Research Institute(本田研究所)
;
Center of Excellence for Computational Cognition, Georgia Tech(佐治亚理工学院计算认知卓越中心)
;
Departments of Neuroscience and Psychology, The University of Texas at Austin(德克萨斯大学奥斯汀分校神经科学和心理学系)
OpenRC: An Open-Source Robotic Colonoscopy Framework for Multimodal Data Acquisition and Autonomy Research
OpenRC:一种用于多模态数据采集和自主性研究的开源机器人结肠镜框架
Siddhartha Kapuria, Mohammad Rafiee Javazm, Naruhiko Ikoma, Joga Ivatury, Mohammad Ali Nasseri, Nassir Navab, Farshid Alambeigi
机构
*
Walker Department of Mechanical Engineering, The University of Texas at Austin(德克萨斯大学奥斯汀分校沃克机械工程系)
;
Department of Surgical Oncology, Division of Surgery, The University of Texas MD Anderson Cancer Center(德克萨斯大学MD安德森癌症中心外科肿瘤学系)
;
School of Medicine and Health, Technical University of Munich(慕尼黑工业大学医学与健康学院)
Large language models reorganize representational geometry during in-context learning
大型语言模型在上下文学习中重组表征几何结构
Hua-Dong Xiong, Li Ji-An, Robert C. Wilson, Kwonjoon Lee, Xue-Xin Wei
机构
*
School of Psychological and Brain Sciences, Georgia Tech(佐治亚理工学院心理与脑科学学院)
;
Department of Psychology, New York University(纽约大学心理学系)
;
Center of Excellence for Computational Cognition, Georgia Tech(佐治亚理工学院计算认知卓越中心)
;
Honda Research Institute(本田研究院)
;
Departments of Neuroscience and Psychology, The University of Texas at Austin(德克萨斯大学奥斯汀分校神经科学与心理学系)
From Words to Widgets for Controllable LLM Generation
从词语到小部件:可控制的LLM生成
Chao Zhang, Yiren Liu, Lunyiu Nie, Jeffrey M. Rzeszotarski, Yun Huang, Tal August
机构
*
Cornell University(康奈尔大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Loyola University Maryland(路易斯安那大学马里兰分校)
When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation
当AI基准测试达到平台期:基准饱和的系统性研究
Mubashara Akhtar, Anka Reuel, Prajna Soni, Sanchit Ahuja, Pawan Sasanka Ammanamanchi, Ruchit Rawal, Vilém Zouhar, Srishti Yadav, Chenxi Whitehouse, Dayeon Ki, Jennifer Mickel, Leshem Choshen, Marek Šuppa, Jan Batzner, Jenny Chim, Jeba Sania, Yanan Long, Hossein A. Rahmani, Christina Knight, Yiyang Nan, Jyoutir Raj, Yu Fan, Shubham Singh, Subramanyam Sahoo, Eliya Habba, Usman Gohar, Siddhesh Pawar, Robert Scholz, Arjun Subramonian, Jingwei Ni, Mykel Kochenderfer, Sanmi Koyejo, Mrinmaya Sachan, Stella Biderman, Zeerak Talat, Avijit Ghosh, Irene Solaiman
机构
*
University of California, Berkeley(加州大学伯克利分校)
;
University of Toronto(多伦多大学)
;
University of Washington(华盛顿大学)
;
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Michigan(密歇根大学)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
机构
*
NVIDIA
;
Georgia Institute of Technology(佐治亚理工学院)
;
Stanford University(斯坦福大学)
;
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
University of Toronto(多伦多大学)
Tianjin Huang, Zhangyang Wang, Haotian Hu, Zhenyu Zhang, Gaojie Jin, Xiang Li, Li Shen, Jiaxing Shang, Tianlong Chen, Ke Li, Lu Liu, Qingsong Wen, Shiwei Liu
机构
*
Department of Computer Science, University of Exeter(埃克塞特大学计算机科学系)
;
Department of Mathematics and Computer Science, Eindhoven University of Technology(埃因霍温理工大学数学与计算机科学系)
;
School of the Gifted Young, University of Science and Technology of China(中国科学技术大学天才青年学院)
;
Department of Electrical and Computer Engineering, University of Texas at Austin(德克萨斯大学奥斯汀分校电气与计算机工程系)
;
Department of Computer Science, University of Reading(阅读大学计算机科学系)
;
School of Cyber Science and Technology, Sun Yat-sen University(中山大学网络科学与技术学院)
;
Department of Computer Science, The University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校计算机科学系)
;
ELLIS Institute Tubingen(图宾根ELLIS研究所)
;
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
;
Tübingen AI Center, Tübingen, Germany(图宾根人工智能中心,德国图宾根)
;
College of Computer Science, Chongqing University(重庆大学计算机学院)
Prior laundering: learned priors with inherited, undetectable overconfidence
先验清洗:具有继承性、不可检测的过度自信的学习先验
Ali Siahkoohi, Sina Alemohammad
机构
*
Institute for Artificial Intelligence, University of Central Florida(人工智能研究所,中央佛罗里达大学)
;
Department of CS, University of Central Florida(计算机科学系,中央佛罗里达大学)
;
Department of ECE, The University of Texas at Austin(电子工程系,德克萨斯大学奥斯汀分校)
COMPOL: A Unified Neural Operator Framework for Scalable Multi-Physics Simulations
COMPOL:用于可扩展多物理场模拟的统一神经算子框架
Junqi Qu, Tao Wang, Yushun Dong, Hewei Tang, Shibo Li
机构
*
School of Information, University of Michigan–Ann Arbor(信息学院,密歇根大学安娜堡分校)
;
Hildebrand Department of Petroleum and Geosystems Engineering, The University of Texas at Austin(石油与地质系统工程系,德克萨斯大学奥斯汀分校)
;
Department of Computer Science, Florida State University(计算机科学系,佛罗里达州立大学)