机构
*
Stanford University(斯坦福大学)
;
Shenzhen University(深圳大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
City University of Hong Kong(香港城市大学)
;
Renmin University of China(中国人民大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);prompting(abstract)
机构
*
The University of New South Wales(新南威尔士大学)
;
The University of Technology Sydney(技术大学悉尼分校)
;
University of Technology Sydney(技术大学悉尼分校)
;
Duke University(杜克大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);pretraining(abstract)
CommentsSurveys and overviews; Natural language processing; Knowledge representation and reasoning; Graph algorithms
Towards Reasoning Ability of Small Language Models
Gaurav Srivastava, Shuxiang Cao, Xuan Wang
机构
*
Department of Computer Science, Virginia Tech, USA(弗吉尼亚理工大学计算机科学系)
;
Department of Physics, Clarendon Laboratory, University of Oxford, UK(牛津大学物理系,克伦德尔实验室)
;
NVIDIA Corporation(英伟达公司)
专题命中
评测与基准
:language model(title,abstract);small language model(title,abstract);LLM(abstract);large language model(abstract)
PortBench: A Correlation-Aware, Full-Pipeline Benchmark for LLM-Driven Portfolio Management
PortBench: 一种相关性感知的、全流水线的LLM驱动投资组合管理基准
Yuxuan Zhao, Sijia Chen, Ningxin Su
机构
*
Yantai Research Institute of Harbin Engineering University(哈尔滨工程大学烟台研究院)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);pretraining(abstract)
ForesightSafety-SAGE:A Fully Automated Scenario Generation and Safety Evaluation Framework for LLM Agents
VESTA: 一种全自动的LLM智能体场景生成与安全评估框架
Lu Jia, Haibo Tong, Feifei Zhao, Jindong Li, Dongqi Liang, Ping Wu, Qian Zhang, Yi Zeng
机构
*
BrainCog AI Lab, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所类脑人工智能实验室)
;
Beijing Institute of AI Safety and Governance (Beijing-AISI)(北京人工智能安全与治理研究院)
;
Beijing Key Laboratory of Safe AI and Superalignment(北京市安全人工智能与超级对齐重点实验室)
;
School of Artificial Intelligence, UCAS(中国科学院大学人工智能学院)
;
Long-term AI(长期人工智能)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI
CommentsThis paper has already been accepted and presented at the IEEE 27th International Conference on Information Reuse and Integration for Data Science (IRI 2026) from July 31 to August 2, 2026, in Seattle, WA, USA