CogAtom: From Cognitive Atoms to Olympiad-level Mathematical Reasoning in Large Language Models
Zhuofan Chen, Jiyuan He, Yichi Zhang, Xing Hu, Haoxing Wen, Jun Bai, Wenge Rong
机构
*
School of Computer Science and Engineering, Beihang University, China(北京航空航天大学计算机科学与工程学院)
;
Meituan Inc.(美团公司)
;
Beijing Institute for General Artificial Intelligence, China(北京通用人工智能研究院)
专题命中
推理与问题求解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG
Do Large Language Models Truly Grasp Mathematics? An Empirical Exploration From Cognitive Psychology
Wei Xie, Shuoyoucheng Ma, Zhenhua Wang, Enze Wang, Kai Chen, Xiaobing Sun, Baosheng Wang
机构
*
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机科学与技术学院)
;
Institute of High Performance Computing, Agency for Science, Technology and Research (A*STAR)(科学、技术和研究局高性能计算研究所)
;
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
专题命中
推理与问题求解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG
CommentsThank you for your attention. This paper was accepted by the CogSci 2025 conference in April and published in August. The location in the proceedings is: https://escholarship.org/uc/item/24x9t7s1
CX-Mind: A Pioneering Multimodal Large Language Model for Interleaved Reasoning in Chest X-ray via Curriculum-Guided Reinforcement Learning
Wenjie Li, Yujie Zhang, Haoran Sun, Yueqi Li, Fanrui Zhang, Mengzhe Xu, Victoria Borja Clausich, Sade Mellin, Renhao Yang, Chenrun Wang, Jethro Zih-Shuo Wang, Shiyi Yao, Gen Li, Yidong Xu, Hanyu Wang, Yilin Huang, Angela Lin Wang, Chen Shi, Yin Zhang, Jianan Guo, Luqi Yang, Renxuan Li, Yang Xu, Jiawei Liu, Yao Zhang, Lei Liu, Carlos Gutiérrez SanRomán, Lei Wang
机构
*
College of Health Science and Technology, Shanghai Jiao Tong University School of Medicine(上海交通大学医学院健康科学与技术学院)
;
Shanghai Innovation Institute(上海创新研究院)
;
Clinical Center for Sports Medicine, Department of Orthopaedics, Ruijin Hospital, Shanghai Jiao Tong University School of Medicine(上海交通大学医学院骨科临床中心)
;
School of Basic Medical Sciences, Intelligent Medicine Institute, Fudan University(复旦大学基础医学学院)
;
Department of Hematology, The First Affiliated Hospital, College of Medicine, Zhejiang University(浙江大学医学院第一附属医院血液科)
;
MoE Key Laboratory of Brain-Inspired Intelligent Perception and Cognition, University of Science and Technology of China(中国科学技术大学脑启发智能感知与认知教育部重点实验室)
;
Department of Public Health and Primary Care, University of Cambridge(剑桥大学公共卫生与初级保健学院)
;
Department of Medicine, Faculty of Health Sciences, Universidad CEU Cardenal Herrera(CEU卡德纳尔-赫尔曼大学健康科学学院医学系)
;
Faculty of Medicine, University of Helsinki(赫尔辛基大学医学院)
;
X-LANCE Lab, School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院X-LANCE实验室)
;
Department of Hepatobiliary Surgery, National Cancer Center / National Clinical Research Center for Cancer / Cancer Hospital, Chinese Academy of Medical Sciences and Peking Union Medical College(中国医学科学院肿瘤医院肝胆外科)
;
Department of Surgery, The Ohio State University Wexner Medical Center, The James Comprehensive Cancer Center(俄亥俄州立大学韦克斯纳医学中心外科部,詹姆斯综合癌症中心)
;
Ningbo Institute of Technology, Beihang University(北航宁波理工学院)
专题命中
推理与问题求解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG
Interpreting Multi-band Galaxy Observations with Large Language Model-Based Agents
Zechang Sun, Yuan-Sen Ting, Yaobo Liang, Nan Duan, Song Huang, Zheng Cai
专题命中
推理与问题求解
:large language model(title,abstract);language model(title,abstract);LLM(abstract)
CommentsAccepted at the NIPS ML4PS Workshop 2024. The journal version is in preparation. Code and data will be fully made public following the journal publication. We welcome any comments and feedback
Integrating External Tools with Large Language Models to Improve Accuracy
Nripesh Niketan, Hadj Batatia
专题命中
推理与问题求解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG
Comments9 pages, 3 figures, 2 tables. Extended version of paper published in Proceedings of International Conference on Information Technology and Applications, Springer Nature Singapore, 2025, pp. 409-421. This version includes additional experimental results comparing against GPT-4o, LLaMA-Large, Mistral-Large, and Phi-Large, expanded evaluation methodology, and enhanced analysis
Journal refProceedings of International Conference on Information Technology and Applications, Springer Nature Singapore, 2025, pp. 409-421
Unveiling Causal Reasoning in Large Language Models: Reality or Mirage?
Haoang Chi, He Li, Wenjing Yang, Feng Liu, Long Lan, Xiaoguang Ren, Tongliang Liu, Bo Han
机构
*
Intelligent Game and Decision Lab(智能游戏与决策实验室)
;
National University of Defense Technology(国防科技大学)
;
University of Melbourne(墨尔本大学)
;
University of Sydney(悉尼大学)
;
Hong Kong Baptist University(香港 Baptist 大学)
专题命中
推理与问题求解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG
Comments24 pages, accepted at NeurIPS 2024
Journal refAdvances in Neural Information Processing Systems, 2024, 37: 96640-96670