SegCompass: Exploring Interpretable Alignment with Sparse Autoencoders for Enhanced Reasoning Segmentation
SegCompass: 探索通过稀疏自编码器实现可解释对齐以增强推理分割
Zhenyu Lu, Liupeng Li, Jinpeng Wang, Haoqian Kang, Yan Feng, Ke Chen, Yaowei Wang
机构
*
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院)
;
Peng Cheng Laboratory(鹏城实验室)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Meituan, Beijing(美团,北京)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
College of Computer Science and Technology, Jilin University(吉林大学计算机科学与技术学院)
Steins;Gate Drive: Semantic Safety Arbitration over Structured Futures for Latency-Decoupled LLM Planning
Steins;Gate Drive: 基于结构化未来语义安全仲裁的延迟解耦LLM规划
Anjie Qiu, Hans D. Schotten
机构
*
Institute for Wireless Communication and Navigation(无线通信与导航研究所)
;
RPTU University Kaiserslautern-Landau(凯撒斯劳滕-兰道大学)
;
German Research Center for Artificial Intelligence(德国人工智能研究中心)
Supervised Classification Heads as Semantic Prototypes: Unlocking Vision-Language Alignment via Weight Recycling
监督分类头作为语义原型:通过权重重用解锁视觉-语言对齐
David Méndez, Roberto Confalonieri, Natalia Díaz Rodríguez
机构
*
Department of Computer Science and Artificial Intelligence, DaSCI Institute, University of Granada, Granada, Spain(计算机科学与人工智能系,DaSCI研究所,格拉纳达大学,格拉纳达,西班牙)
;
Department of Mathematics ``Tullio Levi-Civita'', University of Padova, Padova, Italy(托里利-西维塔数学系,帕多瓦大学,帕多瓦,意大利)
From monoliths to modules: Decomposing transducers for efficient world modelling
从整体到模块:分解转换器以实现高效的world建模
Alexander Boyd, Franz Nowak, David Hyland, Manuel Baltieri, Fernando E. Rosas
机构
*
Department of Informatics, University of Sussex(Sussex大学信息学院)
;
Beyond Institute for Theoretical Science (BITS)(理论科学研究所)
;
ETH Zürich(苏黎世联邦理工学院)
;
Principles of Intelligent Behaviour in Biological and Social Systems (PIBBSS)(生物和社会系统智能行为原理研究所)
;
Department of Computer Science, University of Oxford(牛津大学计算机科学系)
;
Araya Inc.(Araya公司)
;
Sussex AI and Sussex Centre for Consciousness Science, University of Sussex(Sussex大学人工智能与意识科学中心)
;
Centre for Complexity Science and Center for Psychedelic Research, Department of Brain Sciences, Imperial College London(复杂科学中心和迷幻研究中心,伦敦帝国理工学院脑科学系)
;
Center for Eudaimonia and Human Flourishing, University of Oxford(幸福与人类繁荣中心,牛津大学)
机构
*
National Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
School of Intelligent Science and Technology, Nanjing University(南京大学智能科学与技术学院)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
;
Microsoft AI(微软AI)
机构
*
JIUTIAN Research(JIUTIAN研究)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
MIIT Key Laboratory of Data and Decision Intelligence(信息与决策智能重点实验室)
;
Beihang University(北航)
Visibility nowcasting in South Korea: a machine learning approach to class imbalance and distribution shift
韩国可见度现在预测:一种处理数据不平衡和分布偏移的机器学习方法
Bong Gyun Shin, Chan Sik Lee, Hyesun Suh
机构
*
Department of AI Big Data(人工智能大数据系)
;
Daejin University(大 Jain 大学)
;
Department of Statistics and Actuarial Science(统计与精算科学系)
;
Soongsil University(顺斯大学)
;
College of Artificial Intelligence Convergence(人工智能融合学院)
Decoupling Endpoint and Semantic Transition Learning for Zero-Shot Composed Image Retrieval
解耦端点与语义转换学习以实现零样本复合图像检索
Mingyu Liu, Sihan Huang, Yijia Fan, Yinlin Yan, Quan Zhang, Jian-Fang Hu, Jianhuang Lai
机构
*
Sun Yat-sen University(中山大学)
;
Guangdong Province Key Laboratory of Information Security Technology(广东省信息安全技术重点实验室)
;
Key Laboratory of Machine Intelligence and Advanced Computing(机器智能与先进计算重点实验室)
Unveiling the Reasoning Process of Large Language Models
揭示大型语言模型的推理过程
Junjie Zhang, Zhen Shen, Xisong Dong, Gang Xiong
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
Universal Skeleton Understanding via Differentiable Rendering and MLLMs
通过可微渲染和大语言模型实现通用骨架理解
Ziyi Wang, Peiming Li, Xinshun Wang, Yang Tang, Kai-Kuang Ma, Mengyuan Liu
机构
*
State Key Laboratory of General Artificial Intelligence, Peking University, Shenzhen Graduate School, China(人工智能通用基础理论国家重点实验室,北京大学深圳研究生院,中国)
;
Tencent(腾讯)
;
Nanjing University of Aeronautics(南京航空航天大学)
Comments11 pages, 6 figures, 2 tables. Corpus, oracle, output extractor, prompts, harness, self-review probe, and all 1,980 + 1,979 raw model outputs released as supplementary material at https://zenodo.org/records/20300861