Teaching AI Stepwise Diagnostic Reasoning with Report-Guided Chain-of-Thought Learning
Yihong Luo, Wenwu He, Zhuo-Xu Cui, Dong Liang
机构
*
Fujian University of Technology(福建工程学院)
;
Fujian Provincial Key Laboratory of Big Data Mining and Applications(福建省大数据挖掘与应用重点实验室)
;
Key Laboratory of Biomedical Imaging Science and System, Chinese Academy of Sciences(生物医学成像科学与系统重点实验室,中国科学院)
SCOUT: Teaching Pre-trained Language Models to Enhance Reasoning via Flow Chain-of-Thought
Guanghao Li, Wenhao Jiang, Mingfeng Chen, Yan Li, Hao Yu, Shuting Dong, Tao Ren, Ming Tang, Chun Yuan
机构
*
SIGS, Tsinghua University(清华大学信息科学与技术学院)
;
Southern University of Science and Technology(南方科技大学)
;
Guangdong Laboratory of AI and Digital Economy (SZ)(广东人工智能与数字经济实验室)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Peking University(北京大学)
MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs
Jiakang Yuan, Tianshuo Peng, Yilei Jiang, Yiting Lu, Renrui Zhang, Kaituo Feng, Chaoyou Fu, Tao Chen, Lei Bai, Bo Zhang, Xiangyu Yue
机构
*
Fudan University(复旦大学)
;
MMLab, The Chinese University of Hong Kong(中大香港人工智能实验室)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Nanjing University(南京大学)
CommentsJakub Macina, Nico Daheim, and Sankalan Pal Chowdhury contributed equally to this work. Accepted at EMNLP2023 Findings. Code and dataset available: https://github.com/eth-nlped/mathdial
Measuring and curing reasoning rigidity: from decorative chain-of-thought to genuine faithfulness
衡量和治愈推理刚性:从装饰性推理链到真正的忠实
Abhinaba Basu, Pavan Chakraborty
机构
*
Indian Institute of Information Technology Allahabad (IIITA)(印度阿拉哈巴德信息技术学院)
;
National Institute of Electronics and Information Technology (NIELIT)(国家电子与信息技术研究院)
Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization
通过多路径推理和反馈驱动优化实现自动化可视化代码合成
Wonduk Seo, Daye Kang, Hyunjin An, Taehan Kim, Soohyuk Cho, Seungyong Lee, Minhyeong Yu, Jian Park, Yi Bu, Seunghyun Lee
机构
*
AI Research, Enhans, Seoul, South Korea Innovation \& Technology, KAIST, Daejeon, South Korea Department of Computer Science, University of California, Berkeley, CA, United States Department of Electrical
CommentsThis manuscript is withdrawn to allow careful review and correction of bibliographic issues identified after submission, including references that could not be adequately verified. These matters should be resolved before further circulation
Faithful or Just Plausible? Evaluating the Faithfulness of Closed-Source LLMs in Medical Reasoning
忠实还是只是合理?评估闭源LLM在医学推理中的忠实性
Halimat Afolabi, Zainab Afolabi, Elizabeth Friel, Jude Roberts, Antonio Ji-Xu, Lloyd Chen, Egheosa Ogbomo, Emiliomo Imevbore, Phil Eneje, Wissal El Ouahidi, Aaron Sohal, Alisa Kennan, Shreya Srivastava, Anirudh Vairavan, Laura Napitu, Katie McClure
机构
*
Stratified Precision
;
Harvard Medical School(哈佛医学院)
;
Imperial College London(帝国理工学院伦敦分校)
;
National Health Service(国家健康服务系统)
;
Ipsen France(Ipsen法国)
;
University College London(伦敦大学学院)
机构
*
College of Computer Science, Nankai University, Tianjin, China(南开大学计算机科学学院,天津,中国)
;
Academy for Advanced Interdisciplinary Studies, Nankai University, Tianjin, China(南开大学先进跨学科研究学院,天津,中国)
机构
*
School of Data Science, The Chinese University of Hong Kong, Shenzhen (CUHK-Shenzhen)(数据科学学院,香港中文大学(深圳))
;
Faculty of Information Technology, Monash University(信息技术学院,莫纳什大学)
;
Department of Dermatology, Tianjin Institute of Integrative Dermatology, Tianjin Academy of Traditional Chinese Medicine Affiliated Hospital(皮肤科,天津整合皮肤科研究所,天津中医研究院附属医院)
;
Department of Dermatology, The First Affiliated Hospital, Shantou University Medical College(皮肤科,汕头大学医学院第一附属医院)
;
Department of Dermatology, Beijing AnZhen Hospital, Capital Medical University(皮肤科,北京安贞医院,首都医科大学)
;
School of Computing and Data Science, The University of Hong Kong(计算与数据科学学院,香港大学)
;
Institute of Automation, Chinese Academy of Sciences(自动化研究所,中国科学院)
;
Department of Dermatology, Beijing Aerospace General Hospital(皮肤科,北京航天总医院)
机构
*
Institute of Automation, CAS(中国科学院自动化研究所)
;
School of Artifcial Intelligence, UCAS(中国科学技术大学人工智能学院)
;
Beijing Academy of Artificial Intelligence (BAAI)(北京人工智能研究院)
;
Shenzhen International GraduateSchool,Tsinghua University(深圳国际研究生院,清华大学)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,北京大学计算机学院)
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes
LLM推理的周期表:推理范式、方法与失败模式的结构化综述
Avinash Anand, Mahisha Ramesh, Avni Mittal, Ashutosh Kumar, Rishitej Reddy Vyalla, Erik Cambria, Zhengkui Wang, Timothy Liu, Aik Beng Ng, Simon See, Rajiv Ratn Shah
机构
*
Singapore Institute of Technology(新加坡理工大学)
;
Nvidia AI Center (SNAIC)(英伟达人工智能中心(SNAIC))
;
MIDAS Lab, IIIT Delhi(IIIT德里MIDAS实验室)
;
MIDAS Lab, IIT Mandi(IIT曼迪MIDAS实验室)
;
Owl Autonomous Imaging, Inc.(Owl自主成像公司)
;
College of Computing & Data Science, NTU Singapore(新加坡南洋理工大学计算与数据科学学院)
;
NVIDIA AI Technology Centre, Singapore(英伟达新加坡人工智能技术中心)
;
Department of Computer Science and Engineering, IIT Kanpur(IIT坎普尔计算机科学与工程系)