Not All Attention is Needed: Parameter and Computation Efficient Transfer Learning for Multi-modal Large Language Models
并非所有注意力都是必需的:面向多模态大语言模型的参数和计算高效迁移学习
机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China(教育部多媒体可信感知与高效计算重点实验室) ; Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能研究院)
专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.CL
AI总结 本文提出高效注意力跳过方法,通过减少冗余注意力计算提升多模态大语言模型的推理效率与性能。