机构
*
Shanghai Key Lab of Intelligent Information Processing, Fudan University(复旦大学上海智能信息处理重点实验室)
;
School of Computer Science, Fudan University(复旦大学计算机科学技术学院)
;
Shanghai Collaborative Innovation Center of Intelligent Visual Computing(上海智能视觉计算协同创新中心)
;
Youtu Lab, Tencent(腾讯优图实验室)
;
Meta AI
;
Shanghai AI Laboratory(上海人工智能实验室)
CL-CLIP: CLIP-Based Continual Learning Framework with Cost-Volume Category Decoupling for Object Detection
CL-CLIP: 基于CLIP的持续学习框架与代价体积类别解耦用于目标检测
Zihan Liu, Yuguang Yang, Shengjie Su, Jianing Pang, Linlin Yang, Chunyu Xie, Nikolai Yu. Zolotykh, Baochang Zhang
机构
*
National College for Excellent Engineers, Beihang University(卓越工程师学院,北京航空航天大学)
;
AI Research, Qihoo 360(360人工智能研究院,奇虎360)
;
School of Electronic Information Engineering, Beihang University(电子信息学院,北京航空航天大学)
;
School of Cyber Science and Technology, Beihang University(网络安全科学与技术学院,北京航空航天大学)
;
School of Computer Science and Engineering, Beihang University(计算机科学与工程学院,北京航空航天大学)
;
State Key Laboratory of Media Convergence and Communication, Communication University of China(媒体融合与传播国家重点实验室,中国传媒大学)
;
Institute of Information Technology, Mathematics and Mechanics, Lobachebsky University(信息技术、数学与力学学院,洛瓦茨基大学)
;
School of Artificial Intelligence, Beihang University(人工智能学院,北京航空航天大学)
Trade-offs in Medical LLM Adaptation: An Empirical Study in French QA
医学LLM适应中的权衡:法语问答的实证研究
Ikram Belmadani, Oumaima El Khettari, Carlos Ramisch, Frederic Bechet, Richard Dufour, Benoit Favre
机构
*
Aix-Marseille Univ., CNRS, LIS UMR 7020(艾克斯-马赛大学,法国国家科学研究中心,LIS UMR 7020)
;
Nantes Univ., École Centrale Nantes, CNRS, LS2N UMR 6004(南特大学,南特中央理工大学,法国国家科学研究中心,LS2N UMR 6004)
;
Grenoble Alpes Univ., CNRS, INRIA, Grenoble INP, LIG UMR 5217(格勒诺布尔阿尔卑斯大学,法国国家科学研究中心,INRIA,格勒诺布尔INP,LIG UMR 5217)
专题命中
指令微调
:LLM(title,title_cn);SFT(summary_cn,abstract);large language model(abstract);language model(abstract)
Fine-Tuning Large Language Models for Quantum Reasoning
微调大型语言模型用于量子推理
Katherine Ip, Casey R. Myers, Udaya Parampalli, James Quach, Peiyong Wang
机构
*
School of Computing and Information Systems, The University of Melbourne(墨尔本大学计算机与信息系统学院)
;
School of Physics, Mathematics and Computing, The University of Western Australia(西澳大学物理、数学与计算学院)
;
Pawsey Supercomputing Centre(帕韦西超级计算中心)
;
CSIRO Clayton(CSIRO 克莱顿分校)
专题命中
指令微调
:SFT(summary_cn,abstract);large language model(title,abstract);language model(title,abstract);LLM(summary_cn,abstract_cn)
When RL Fails after SFT: Rejuvenating Model Plasticity for Robust SFT-to-RL Handoff
当强化学习在监督微调后失效:恢复模型可塑性以实现稳健的SFT到RL交接
Runze Liu, Jiashun Liu, Xu Wan, Yuqian Fu, Ling Pan
机构
*
Hong Kong University of Science and Technology(香港科技大学)
;
Zhejiang University(浙江大学)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA(多模态人工智能系统国家重点实验室,CASIA)
专题命中
指令微调
:SFT(title,title_cn);LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)