LiteLMGuard: Seamless and Lightweight On-Device Prompt Filtering for Safeguarding Small Language Models against Quantization-induced Risks and Vulnerabilities
机构
*
Tsinghua University(清华大学)
;
Infinigence AI
;
The Chinese University of Hong Kong(香港中文大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
SLAI
;
Shanghai AI Laboratory(上海人工智能实验室)
专题命中
效率与部署
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.LG
Speech recognition assisted by large language models to command software orally -- Application to an augmented and virtual reality web app for immersive molecular graphics
通过大型语言模型辅助的语音识别来控制软件——应用于增强和虚拟现实网页应用中的沉浸式分子图形
Fabio Cortes Rodriguez, Luciano Abriata
专题命中
效率与部署
:large language model(title,abstract);language model(title,abstract);LLM(abstract)
Comments14 pages, 13 figures, 6 tables, 7 algorithms, 16 references, submitted to ACM/IEEE International Conference on Systems and Software Engineering
Bridging the Gap Between Promise and Performance for Microscaling FP4 Quantization
弥合承诺与性能之间的差距:为微缩FP4量化 bridging the gap between promise and performance for microscaling FP4 quantization
Vage Egiazarian, Roberto L. Castro, Denis Kuznedelev, Andrei Panferov, Eldar Kurtic, Shubhra Pandit, Alexandre Marques, Mark Kurtz, Saleh Ashkboos, Torsten Hoefler, Dan Alistarh
机构
*
Institute of Science and Technology Austria(奥地利科学与技术研究院)
;
Yandex Research(Yandex研究)
;
Red Hat AI(红帽人工智能)
;
ETH Zürich(苏黎世联邦理工学院)
专题命中
效率与部署
:LLM(abstract);large language model(abstract);language model(abstract);post-training(abstract)
SUN: Shared Use of Next-token Prediction for Efficient Multi-LLM Disaggregated Serving
SUN: 为高效多LLM解耦服务实现下一个token预测的共享使用
Sunghyeon Woo, Ahreum Seo, Jaegwang Lee, Jaeeun Kil, Hanbae Seo, Joonghoon Kim, Baeseong Park, Se Jung Kwon, Dongsoo Lee
机构
*
Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家)
;
School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家)
Generalizable Knowledge Distillation from Vision Foundation Models for Semantic Segmentation
从视觉基础模型中实现可泛化的知识蒸馏用于语义分割
Chonghua Lv, Dong Zhao, Shuang Wang, Dou Quan, Ning Huyan, Nicu Sebe, Zhun Zhong
机构
*
School of Artificial Intelligence, Xidian University, China(西安电子科技大学人工智能学院)
;
Department of Information Engineering and Computer Science, University of Trento, Italy(特伦托大学信息工程与计算机科学系)
;
Department of Automation, Tsinghua University, China(清华大学自动化系)
;
School of Computer Science and Information Engineering, Hefei University of Technology, China(合肥工业大学计算机科学与信息工程学院)
机构
*
THU(清华大学)
;
USTC(University of Science and Technology of China)
;
BUAA(Beijing University of Aeronautics and Automation)
;
PKU(Peking University)
;
TJU(Tianjin University)
专题命中
效率与部署
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
Agentic Self-Evolutionary Replanning for Embodied Navigation
代理自进化重规划用于具身导航
Guoliang Li, Ruihua Han, Chengyang Li, He Li, Shuai Wang, Wenchao Ding, Hong Zhang, Chengzhong Xu
机构
*
Department of Computer and Information Science, University of Macau(澳门大学计算机与信息科学系)
;
SIAT, Chinese Academy of Sciences(中国科学院上海技术物理研究所)
;
Department of Computer Science, University of Hong Kong(香港大学计算机科学系)
;
Academy for Engineering & Technology, Fudan University(复旦大学工程与技术学院)
;
Department of EEE, Southern University of Science and Technology(南方科技大学电子工程系)
专题命中
效率与部署
:LLM(abstract);large language model(abstract);language model(abstract)
AI总结
SERP通过自进化动作模型和图链式思考重规划,提升具身导航在复杂环境中的鲁棒性和效率。
Comments8 pages, 10 figures, 4 tables, submitted to IEEE for possible publication