Beyond Memory Leaderboards: Evaluating Scientific Memory as Budgeted Context Restoration
超越内存排行榜:将科学记忆评估为预算上下文恢复
Maksim Sheverev, David Finkelstein, Sergey Nikolenko
机构
*
Quantellence Research(Quantellence研究公司)
;
St. Petersburg Department of the Steklov Institute of Mathematics(斯捷克洛夫数学研究所圣彼得堡分部)
;
St. Petersburg State University(圣彼得堡国立大学)
机构
*
VNU University of Engineering and Technology(越南国立工程技术大学)
;
Center for Juris-Informatics, ROIS-DS College of Engineering & Computer Science, VinUniversity(越南Vin大学工程与计算机科学学院法律信息学中心)
Closing the Prior-Posterior Loop: Self-Reflective Molecular Design with Analysis-Driven LLM Iteration
闭合先验-后验循环:基于分析驱动LLM迭代的自反性分子设计
Junyi Gong, Zijie Qiu, Ben Zhong Tang
机构
*
Faculty of Chemistry, Shenzhen MSU-BIT University(深圳MSU-BIT大学化学学院)
;
School of Science and Engineering, Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳)科学与工程学院)
;
Department of Chemistry, Hong Kong University of Science and Technology(香港科技大学化学系)
A Multi-Agent Framework for Zero-Dimensional Reduced-Order Model Planning
用于零维降阶模型规划的多智能体框架
Bingteng Sun, Hao Yin, Yiling Chen, Renjie Xiao, Lei Xie, Shanyou Wang, Ruonan Wang, Shubao Chen, Qingzong Xu, Lin Lu, Qiang Du, Junqiang Zhu
机构
*
Institute of Engineering Thermophysics, Chinese Academy of Sciences(中国科学院工程热物理研究所)
;
National Key Laboratory of Science and Technology on Advanced Light-duty Gas-turbine(先进轻型燃气轮机技术重点实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Qingdao Institute of Aeronautical Technology(青岛航空技术研究院)
;
Nanjing Future Energy System Research Institute(南京未来能源系统研究院)
MMAgent-R$^2$: Learning to Rerank and Reject for Agentic mRAG
MMAgent-R$^2$:用于智能mRAG的重排与拒绝学习
Tao Zhang, Ziqi Zhang, Zongyang Ma, Yuxin Yang, Bing Li, Chunfeng Yuan, Kang Rong, Fengyun Rao, Jing Lyu, Weiming Hu
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA(中国科学院复杂系统管理与控制国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京多模态信息超智能安全重点实验室)
;
WeChat Vision, Tencent Inc.(腾讯微信视觉团队)
;
PeopleAI Inc.(人智公司)
;
School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)
ClinRAG-GRAPH: Clinical-prior Retrieval-Augmented Graph Model with Domain Adversarial Learning for Breast pCR Prediction
ClinRAG-GRAPH: 基于临床先验的检索增强图模型与领域对抗学习用于乳腺癌pCR预测
Yaofei Duan, Yuhao Huang, Tianyu Zhang, Yuan Gao, Luyi Han, Xin Wang, Xinyu Xie, Xinglong Liang, Chunyao Lu, Muzhen He, Patrick Pang, Yue Sun, Ning Mao, Tao Tan, Ritse Mann
机构
*
Department of Medical Imaging, Radboud University Medical Center, Nijmegen, The Netherlands(拉德堡德大学医学中心医学影像科,奈梅亨,荷兰)
;
The Netherlands Cancer Institute, Amsterdam, The Netherlands(荷兰癌症研究所,阿姆斯特丹,荷兰)
;
Faculty of Applied Sciences, Macao Polytechnic University, Macau, China(澳门理工大学应用科学学院,澳门,中国)
;
Boston Children’s Hospital, Harvard Medical School, Boston, USA(波士顿儿童医院,哈佛医学院,波士顿,美国)
;
Centre for Artificial Intelligence and Robotics, Hong Kong Institute of Science & Innovation, Chinese Academy of Sciences, Hong Kong, China(中国科学院香港创新研究院人工智能与机器人创新中心,香港,中国)
;
Imaging Division, University Medical Center Utrecht, Utrecht, The Netherlands(乌得勒支大学医学中心影像部,乌得勒支,荷兰)
;
Department of Radiology, Fuzhou University Affiliated Provincial Hospital, Fuzhou, China(福州大学附属省立医院放射科,福州,中国)
;
Department of Radiology, Yantai Yuhuangding Hospital, Shandong, China(烟台毓璜顶医院放射科,山东,中国)
LODESTAR: Trustworthy Entropy Is Navigated, Not Merely Measured -- Reinforced Polarizer Keeps a Frozen LLM from Being Confidently Misled by the Wrong Evidence
Controlled Memory Interference in Continual LLM Agents
持续大型语言模型智能体中的可控记忆干扰
Ao Ding, Hongzong LI, Shiqin Tang, Li Zhang, Liang Chen, Xuyang Chen, Zi Liang
机构
*
China University of Geosciences (Beijing)(中国地质大学(北京))
;
Northwestern Polytechnical University(西北工业大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Chinese Academy of Sciences(中国科学院)
;
The Hong Kong Polytechnic University(香港理工大学)
;
École polytechnique fédérale de Lausanne(洛桑联邦理工学院)
;
National University of Singapore(新加坡国立大学)
Simulator-Grounded Large Language Models for Industrial Causal Reasoning: Tool-Use, Structured Injection, and Plant-Portable Retrieval for Wastewater Treatment Decision Support