Test-time Offline Reinforcement Learning on Goal-related Experience
测试时的离线强化学习在目标相关经验上的应用
Marco Bagatella, Mert Albaba, Jonas Hübotter, Georg Martius, Andreas Krause
机构
*
ETH Zurich, Zurich, Switzerland(苏黎世联邦理工学院,苏黎世,瑞士)
;
Max Planck Institute for Intelligent Systems, Tubingen, Germany(智能系统马克斯·普朗克研究所,图宾根,德国)
;
University of Tubingen, Tubingen, Germany(图宾根大学,图宾根,德国)
The Expressivity Boundary of Probabilistic Circuits: A Comparison with Large Language Models
概率电路的表达性边界:与大语言模型的比较
Zhiyu Zhao, Xuejie Liu, Muhan Zhang, Anji Liu
机构
*
School of Computing, National University of Singapore(新加坡国立大学计算机学院)
;
Institute for Artificial Intelligence, Peking University(北京大学人工智能研究院)
;
School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学院)
专题命中
效率与部署
:large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG
OmniSIFT: Modality-Asymmetric Token Compression for Efficient Omni-modal Large Language Models
OmniSIFT: 多模态非对称令牌压缩用于高效的多模态大语言模型
Yue Ding, Yiyan Ji, Jungang Li, Xuyang Liu, Xinlong Chen, Junfei Wu, Bozhou Li, Bohan Zeng, Yang Shi, Yushuo Guan, Yuanxing Zhang, Jiaheng Liu, Qiang Liu, Pengfei Wan, Liang Wang
机构
*
New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences (CASIA)(模式识别新实验室(NLPR)、自动化研究所、中国科学院(CASIA))
;
Nanjing University(南京大学)
;
The Hong Kong University of Science(香港科学大学)
;
Sichuan University(四川大学)
;
Peking University(北京大学)
专题命中
效率与部署
:large language model(title,abstract);language model(title,abstract);分类 cs.CL
FOAM: Blocked State Folding for Memory-Efficient LLM Training
FOAM:用于内存高效大语言模型训练的阻塞状态折叠
Ziqing Wen, Jiahuan Wang, Ping Luo, Dongsheng Li, Tao Sun
机构
*
National Key Laboratory of Parallel and Distributed Computing(并行与分布式计算国家重点实验室)
;
College of Computer Science and Technology(计算机科学与技术学院)
;
National University of Defense Technology(国防科技大学)
专题命中
效率与部署
:LLM(title);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Comments9 pages, 16 figures. Accepted at the ICLR 2026 Workshop on Principled Design for Trustworthy AI: Interpretability, Robustness, and Safety across Modalities
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Shanghai Jiao Tong University(上海交通大学)
Jan Arne Telle, Brigt Håvardstun, Jose Hernandez-Orallo
机构
*
Department of Informatics University of Bergen(卑尔根大学信息学院)
;
University of Bergen(卑尔根大学)
;
VRAIN - Universitat Politecnica de Valencia(瓦伦西亚理工大学VRAIN实验室)
;
Universitat Politecnica de Valencia(瓦伦西亚理工大学)
;
Leverhulme Centre for the Future of Intelligence - University of Cambridge(剑桥大学未来智能中心)
专题命中
效率与部署
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG
All Circuits Lead to Rome: Rethinking Functional Anisotropy in Circuit and Sheaf Discovery for LLMs
所有电路都通向罗马:重新思考用于大语言模型的电路和sheaf发现中的功能性各向异性
Xi Chen, Mingyu Jin, Jingcheng Niu, Yutong Yin, Jinman Zhao, Bangwei Guo, Dimitris N. Metaxas, Zhaoran Wang, Yutao Yue, Gerald Penn
机构
*
Rutgers University(罗格斯大学)
;
Northwestern University(西北大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
University of Toronto(多伦多大学)
专题命中
效率与部署
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL
EDITS: Enhancing Dataset Distillation with Implicit Textual Semantics
EDITS: 通过隐式文本语义增强数据集蒸馏
Qianxin Xia, Jiawei Du, Guoming Lu, Zhiyong Shu, Jielei Wang
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Centre for Frontier AI Research, Agency for Science, Technology and Research(前沿人工智能研究中心,科技研究局)
专题命中
效率与部署
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)
机构
*
Institute for AI Industry Research (AIR), Tsinghua University(人工智能产业研究院(AIR),清华大学)
;
Juhaokan Technology Co.,Ltd(极皓科技有限公司)
;
Nanjing University(南京大学)
;
University of Science and Technology of China(中国科学技术大学)