TS-MLLM: A Multi-Modal Large Language Model-based Framework for Industrial Time-Series Big Data Analysis
TS-MLLM:一种基于多模态大语言模型的工业时间序列大数据分析框架
Haiteng Wang, Yikang Li, Yunfei Zhu, Jingheng Yan, Lei Ren, Laurence T. Yang
机构
*
School of Automation Science and Electrical Engineering, Beihang University(北京航空航天大学自动化科学与电气工程学院)
;
School of Software, Beihang University(北京航空航天大学软件学院)
;
Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院)
;
State Key Laboratory of Intelligent Manufacturing System Technology(智能制造系统技术国家重点实验室)
;
School of Computer and Artificial Intelligence, Zhengzhou University(郑州大学计算机与人工智能学院)
;
Department of Computer Science, St. Francis Xavier University(圣弗朗西斯科大学计算机科学系)
机构
*
James Watt School of Engineering, University of Glasgow(格拉斯哥大学詹姆斯·瓦特工程学院)
;
University of Bristol(布里斯托大学)
;
School of Engineering Mathematics and Technology, University of Bristol(布里斯托大学工程数学与技术学院)
;
School of Computing Science, University of Glasgow(格拉斯哥大学计算科学学院)
;
Department of Civil, Environmental & Geomatic Engineering, University College London (UCL)(伦敦大学学院(UCL)土木、环境与测绘工程系)
Evolving Safety Landscape of Multi-modal Large Language Models: A Survey of Emerging Threats and Safeguards
多模态大语言模型的演进安全态势:新兴威胁与防护措施综述
Xi Li, Shu Zhao, Xiaohan Zou, Fei Zhao, Fuxiao Liu, Yusen Zhang, Cheng Han, Yushun Dong, Jiaqi Wang
机构
*
University of Alabama at Birmingham(阿拉巴马大学伯明翰分校)
;
NVIDIA(英伟达公司)
;
Penn State University(宾夕法尼亚州立大学)
;
Columbia University(哥伦比亚大学)
;
University of Missouri-Kansas City(密苏里大学堪萨斯分校)
;
Florida State University(佛罗里达州立大学)
;
Auburn University(奥本大学)
Seeing Before Reasoning: Decoupling Perception and Reasoning for Shortcut-Resilient Multimodal On-Policy Self-Distillation
先看后思:解耦感知与推理以实现抗捷径的多模态在策略自蒸馏
Sihan Wang, Xiyao Liu, Lianqing Liu, Zhi Han
机构
*
State Key Laboratory of Robotics and Intelligent Systems, Shenyang Institute of Automation, Chinese Academy of Sciences(机器人与智能系统国家重点实验室,沈阳自动化研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding
PAR3D: 一种用于场景理解的统一部件感知3D多模态大语言模型
Shaohui Dai, Yansong Qu, You Shen, Shengchuan Zhang, Liujuan Cao
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(教育部多媒体可信感知与高效计算重点实验室,厦门大学)
Chronicle: A Multimodal Foundation Model for Joint Language and Time Series Understanding
Chronicle:一种用于联合语言和时间序列理解的多模态基础模型
Paul Quinlan, Jeremy Levasseur, Qingguo Li, Xiaodan Zhu
机构
*
InertialAI
;
Department of Electrical and Computer Engineering, Queen’s University(皇后大学电气与计算机工程系)
;
Department of Mechanical and Materials Engineering, Queen’s University(皇后大学机械与材料工程系)
专题命中
多模态训练与对齐
:multimodal(title,abstract);multimodal foundation model(title);cross-modal(abstract);分类 cs.CL、cs.AI
Cross-modal Affinity-aligned Multimodal Learning Analytics for Predicting Student Collaboration Satisfaction in Game-Based Learning
跨模态亲和对齐的多模态学习分析用于预测基于游戏的学习中学生协作满意度
Wen-Hsin Tsai, Chia-Ming Lee, Yuk-Ying Tung
机构
*
Institute of Education, National Cheng Kung University(国立成功大学教育研究所)
;
Institute of Intelligent System, National Yang Ming Chiao Tung University(阳明交通大学智能系统研究所)
;
Department of Computer Science, University at Albany, State University of New York(纽约州立大学水牛城分校计算机科学系)
Dynamic Cross-Modal Prompt Generation for Multimodal Continual Instruction Tuning
动态跨模态提示生成用于多模态持续指令微调
Tao Hu, Da-Wei Zhou
机构
*
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
;
State Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
机构
*
Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative, Institute of Digital Twin, EIT, Ningbo(宁波空间智能与数字衍生关键实验室,数字孪生研究院,EIT,宁波)
;
Shanghai Jiao Tong University(上海交通大学)
;
Hong Kong Polytechnic University(香港理工大学)
;
Meituan Inc.(美团公司)
;
National University of Singapore(新加坡国立大学)