Distributional Soft Bellman Operator under the Cramér Geometry
在克拉默几何下的分布软贝尔曼算子
Keru Wang, Yixin Deng, Yao Lyu, Stephen Redmond, Shengbo Eben Li
机构
*
School of Electrical and Electronic Engineering, University College Dublin(都柏林大学学院电气与电子工程学院)
;
School of Vehicle and Mobility, Tsinghua University(清华大学车辆与运载学院)
;
College of Artificial Intelligence, Tsinghua University(清华大学人工智能学院)
机构
*
School of Environment, Tsinghua University(清华大学环境学院)
;
College of Economics and Management, Beijing University of Technology(北京工业大学经济与管理学院)
;
State Key Laboratory of Iron and Steel Industry Environmental Protection, School of Environment, Tsinghua University(清华大学环境学院钢铁工业环境保护国家重点实验室)
;
Appraisal Center for Environmental Engineering, Ministry of Ecology and Environment(生态环境部环境工程评估中心)
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
Nankai University(南开大学)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院)
;
Tsinghua University(清华大学)
;
Harbin Institute of Technology(哈尔滨工业大学)
DepthART: Scaling Foundation Monocular Depth to Tiny Models
DepthART:将基础单目深度模型扩展到小型模型
Feng Xue, Wu Chen, Mingshuai Zhao, Guofeng Zhong, Anlong Ming, Haozhe Wang, Dianqiao Lei, Zhaowen Lin, Haiyang Zhang, Nicu Sebe
机构
*
University of Trento(特伦托大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Tsinghua University(清华大学)
Fourier Geometric Wind Power Forecasting with Numerical Weather Prediction
基于数值天气预报的傅里叶几何风力发电预测
Shiyuan Piao, Fan Zehui, Yang Liu, Hong Cheng, Juepeng Zheng, Jie Zhou, Fugee Tsung
机构
*
The Hong Kong University of Science and Technology(香港科技大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Tsinghua University(清华大学)
;
Goldwind Science and Technology Co.,Ltd(金风科技股份有限公司)
Look Clearly Before Answering: Mitigating Hallucinations in LVLMs via Saliency-Driven Perceptual Realignment
回答前看清楚:通过显著性驱动的感知重新对齐减轻LVLMs中的幻觉
Pengxu Chen, Yao Zhu, Guangming Zhu, Jun Sheng, Jincai Huang, Xiangyang Ji, Liang Zhang
机构
*
Xidian University(西安电子科技大学)
;
Tsinghua University(清华大学)
;
Shanghai Road Transport Development Center(上海市道路运输发展中心)
;
Hunan Institute of Advanced Technology(湖南先进技术研究院)
机构
*
Institute of Artificial Intelligence, State Key Laboratory of Virtual Reality Technology and Systems, Beihang University(北京航空航天大学虚拟现实技术与系统国家重点实验室人工智能研究院)
;
College of Artificial Intelligence, Tsinghua University(清华大学人工智能学院)
;
Security Department, Alibaba Group(阿里巴巴集团安全部)
;
School of Cyber Science and Technology, Shenzhen Campus of Sun Yat-Sen University(中山大学深圳校区网络科学与技术学院)
Multi-level context Modeling for consistent expert selection in Mixture-of-Experts
用于混合专家模型中一致专家选择的多层次上下文建模
Shuhan Huang, Naifan Zhang, Yuanbo Tang, Yang Li, Wai Kin Victor Chan
机构
*
Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)
;
School of AI, The Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳)人工智能学院)
GenHOI: Contact-Aware Humanoid-Object Interaction by Imitating Generated Videos without Task-Specific Training
GenHOI: 通过模仿生成视频实现接触感知的人形机器人-物体交互,无需任务特定训练
Zhihai Bi, Qiang Zhang, Guoyang Zhao, Jiahang Cao, Xueyin Luo, Yushan Zhang, Jinglan Xu, Ruoyu Geng, Yulin Li, Andrew F. Luo, Jun Ma
机构
*
The University of Tokyo(东京大学)
;
National University of Singapore(新加坡国立大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Tsinghua University(清华大学)
机构
*
Independent Researchers(独立研究者)
;
Tencent(腾讯)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Tsinghua University(清华大学)
GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning
GradAlign: 用于大语言模型强化学习的梯度对齐数据选择
Ningyuan Yang, Weihua Du, Weiwei Sun, Sean Welleck, Yiming Yang
机构
*
Institute for Interdisciplinary Information Sciences (IIIS), Tsinghua University(交叉信息学院(IIIS)、清华大学)
;
Language Technologies Institute (LTI), Carnegie Mellon University(语言技术研究所(LTI)、卡内基梅隆大学)
MMR-V: What's Left Unsaid? A Benchmark for Multimodal Deep Reasoning in Videos
MMR-V:未言明的是什么?视频中多模态深度推理的基准测试
Kejian Zhu, Zhuoran Jin, Hongbang Yuan, Jiachun Li, Shangqing Tu, Pengfei Cao, Yubo Chen, Kang Liu, Jun Zhao
机构
*
The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences, Beijing, China(认知与决策智能复杂系统重点实验室,自动化研究所,中国科学院,北京,中国)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)
;
Tsinghua University(清华大学)
Unsupervised Multimodal Clustering for Semantics Discovery in Multimodal Utterances
用于多模态话语语义发现的无监督多模态聚类
Hanlei Zhang, Hua Xu, Fei Long, Xin Wang, Kai Gao
机构
*
State Key Laboratory of Intelligent Technology and Systems, Department of Computer Science and Technology, Tsinghua University(智能技术与系统国家重点实验室,计算机科学与技术系,清华大学)
;
School of Information Science and Engineering, Hebei University of Science and Technology(信息科学与工程学院,河北科技大学)
;
Samton (Jiangxi) Technology Development Co.,Ltd(江西松通科技发展有限公司)