ZeroWBC: Learning Natural Whole-Body Humanoid Interaction from Human Egocentric Data
ZeroWBC: 从人类自我中心数据学习自然全身人形交互
Haoran Yang, Jiacheng Bao, Yucheng Xin, Haoming Song, Yuyang Tian, Bin Zhao, Dong Wang, Xuelong Li
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Northwestern Polytechnical University(西北工业大学)
;
Tsinghua University(清华大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
TeleAI, China Telecom(TeleAI,中国电信)
Positive Alignment: Artificial Intelligence for Human Flourishing
积极对齐:人工智能促进人类繁荣
Ruben Laukkonen, Seb Krier, Chloé Bakalar, Shamil Chandaria, Morten Kringelbach, Adam Elwood, Daniel Ford, Fernando Rosas, Maty Bohacek, Matija Franklin, Nenad Tomašev, Stephanie Chan, Verena Rieser, Roma Patel, Michael Levin, Arun Rao
机构
*
Department of Psychiatry, University of Oxford(牛津大学精神病学系)
;
Flourishing Intelligence Program, Centre for Eudaimonia and Human Flourishing, Linacre College, University of Oxford(牛津大学幸福智能计划、幸福与人类繁荣中心、林acre学院)
;
Google DeepMind(谷歌DeepMind)
;
LIFE
;
OpenAI
;
Anthropic
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Aily Labs(Aily实验室)
;
Stanford University(斯坦福大学)
;
Tufts University(塔夫茨大学)
;
Positive AI Labs(积极AI实验室)
;
Department of Informatics, University of Sussex(Sussex大学信息学系)
;
Department of Brain Sciences, Imperial College London(伦敦帝国理工学院脑科学系)
Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs
面向高效视频大语言模型的sink-token感知剪枝:用于细粒度视频理解
Kibum Kim, Jiwan Kim, Kyle Min, Yueqi Wang, Jinyoung Moon, Julian McAuley, Chanyoung Park
机构
*
Korea Advanced Institute of Science and Technology (KAIST)(韩国高级科学技术研究院)
;
Oracle
;
University of California, San Diego(加州大学圣地亚哥分校)
;
Electronics and Telecommunications Research Institute (ETRI)(电子电信研究院)
When Safe Concepts Become Unsafe: Multi-Concept Compositional Vulnerabilities in Text-to-Image Models
TwoHamsters:文本到图像模型中多概念组合不安全性的基准测试
Chaoshuo Zhang, Yibo Liang, Mengke Tian, Chenhao Lin, Zhengyu Zhao, Le Yang, Chong Zhang, Yang Zhang, Qian Wang, Chao Shen
机构
*
School of Cyber Science and Engineering, Xi'an Jiaotong University(西安交通大学计算机科学与工程学院)
;
CISPA Helmholtz Center for Information Security(信息安全研究中心)
;
School of Cyber Science and Engineering, Wuhan University(武汉大学计算机科学与工程学院)
Co-policy: Responsive Human-Robot Co-Creation for Musical Performances
Co-policy: 响应式人机音乐共创框架
Xuetao Li, Wenke Huang, Mang Ye, Zijian Liu, Jinhua Xie, Jifeng Xuan, Miao Li
机构
*
School of Computer Science, Wuhan University(武汉大学计算机学院)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
School of Automation, Wuhan University of Technology(武汉理工大学自动化学院)
;
School of Geodesy and Geomatics, Wuhan University(武汉大学测绘学院)
;
School of Robotics, Wuhan University(武汉大学机器人学院)
G-IdiomAlign: A Gloss-Pivoted Benchmark for Cross-Lingual Idiom Alignment
G-IdiomAlign:基于释义的跨语言习语对齐基准
Fengying Ye, Yanming Sun, Runzhe Zhan, Zheqi Zhang, Lidia S. Chao, Derek F. Wong
机构
*
NLP 2 CT Lab, Department of Computer and Information Science, University of Macau(NLP 2 CT实验室,计算机与信息科学系,澳门大学)
;
Faculty of Arts and Humanities, University of Macau(人文学院,澳门大学)
iTRIALSPACE: Programmable Virtual Lesion Trials for Controlled Evaluation of Lung CT Models
iTRIALSPACE:用于肺CT模型受控评估的可编程虚拟病灶试验
Fakrul Islam Tushar, Umme Hafsa Momy, Joseph Y. Lo, Geoffrey D. Rubin
机构
*
Department of Radiology and Imaging Sciences, University of Arizona(亚利桑那大学放射科和影像科学系)
;
Department of Biomedical Engineering, Florida International University(佛罗里达国际大学生物医学工程系)
;
Center for Virtual Imaging Trials, Department of Radiology, Duke University Medical Center(达特茅斯大学医学中心虚拟成像试验中心,放射科)
Attention, not scale, drives human-AI alignment in multimodal language prediction
注意力,而非规模,驱动多模态语言预测中的人机对齐
Viktor Kewenig, Andrew Lampinen, Samuel A. Nastase, Christopher Edwards, Quitterie Lacome D'Elascombe, Akilles Rechardt, Jeremy I Skipper, Gabriella Vigliocco
机构
*
Psychology and Language Science, Experimental Psychology, University College London, London, UK(心理学与语言科学、实验心理学,伦敦大学学院,伦敦,英国)
;
Google Deepmind, Mountain View, US(谷歌DeepMind,山景城,美国)
;
Princeton Neuroscience Institute, Princeton University, Princeton, NJ, USA(普林斯顿神经科学研究所,普林斯顿大学,普林斯顿,新泽西州,美国)
;
Computer Science Department, Exeter University(计算机科学系,埃克塞特大学)
机构
*
Beijing Key Laboratory of Intelligent Information Technology, School of Computer Science & Technology, Beijing Institute of Technology(北京智能信息科技重点实验室,计算机科学与技术学院,北京理工大学)
;
State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI)
;
State Key Laboratory of General Artificial Intelligence, Peking University(通用人工智能国家重点实验室,北京大学)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
Guangdong Laboratory of Machine Perception and Intelligent Computing, Shenzhen MSU-BIT University(广东机器感知与智能计算实验室,深圳MSU-BIT大学)
;
Department of Automation, Tsinghua University(自动化系,清华大学)
专题命中
VLM训练与架构
:vision language model(abstract);分类 cs.CV