Comments7 pages, 3 figures, In Proceedings of the 1st ACM SIGPLAN International Workshop on Language Models and Programming Languages (LMPL'25), October 12-18, 2025, Singapore, Singapore. ACM, New York, NY, USA
Mind the Gap: Evaluating the Representativeness of Quantitative Medical Language Reasoning LLM Benchmarks for African Disease Burdens
Fred Mutisya, Shikoh Gitau, Christine Syovata, Diana Oigara, Ibrahim Matende, Muna Aden, Munira Ali, Ryan Nyotu, Diana Marion, Job Nyangena, Nasubo Ongoma, Keith Mbae, Elizabeth Wamicha, Eric Mibuari, Jean Philbert Nsengemana, Talkmore Chidede
专题命中
代码与定理证明
:reasoning(title);分类 cs.AI
CommentsPreprint. 26 pages, includes appendix and tables
Why We Need World Models for AGI: Where LLMs Fail and How World Models May Outperform
为什么我们需要世界模型来实现通用人工智能:大语言模型失败之处以及世界模型如何可能超越
Feisal Alaswad, Batoul Aljaddouh, Maher Alrahhal, Poovammal E, Talal Bonny
机构
*
Department of Computing Technologies(计算技术系)
;
SRM Institute of Science and Technology(SRM科学与技术学院)
;
Bio-Sensing and Bio-Sensors Group(生物传感与生物传感器组)
;
Smart Automation and Communication Technologies Research Institute of Sciences and Engineering(科学与工程智能自动化与通信技术研究所)
;
University of Sharjah, UAE(阿联酋沙迦大学)
;
Department of Computer Engineering(计算机工程系)
;
College of Computing and Informatics(计算与信息学院)
机构
*
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院)
CommentsExperimental verification and more formal argument for Markov approximation of bias propagation to be released soon. Primarily pushed now to establish novelty and ease of sharing. Please do not cite this work until the forthcoming experimental validation and updated mathematical model are provided
Shrinking the Generation-Verification Gap with Weak Verifiers
缩小生成-验证差距的弱验证器
Jon Saad-Falcon, E. Kelly Buchanan, Mayee F. Chen, Tzu-Heng Huang, Brendan McLaughlin, Tanvir Bhathal, Shang Zhu, Ben Athiwaratkun, Frederic Sala, Scott Linderman, Azalia Mirhoseini, Christopher Ré
机构
*
Stanford University(斯坦福大学)
;
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
;
Together AI