arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Beihang University(北京航空航天大学)

2026-04-28 至 2026-04-28 共收录 10
2604.24396 2026-04-28 cs.CV cs.AI

Global Context or Local Detail? Adaptive Visual Grounding for Hallucination Mitigation

全局上下文还是局部细节?面向幻觉缓解的自适应视觉 grounding

Yubo Jiang, Xin Yang, Abudukelimu Wuerkaixi, Zheming Yuan, Xuxin Cheng, Fengying Xie, Zhiguo Jiang, Cao Liu, Ke Zeng, Haopeng Zhang

机构 * School of Astronautics, Beihang University(北航航天学院) Longcat Interaction Team, Meituan(美团Longcat交互团队) Tianmushan Laboratory, Beihang University(北航天门山实验室)

AI总结 本文提出PND框架,通过双路径对比在解码过程中增强视觉真实性,减少幻觉并提升描述细节,无需模型微调。

Comments 9 pages, 8 figures, Findings of ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12009 2026-04-28 cs.CV

LatentStealth: Unnoticeable and Efficient Adversarial Attacks on Expressive Human Pose and Shape Estimation

LatentStealth: 一种无痕且高效的对抗攻击方法用于表达性人体姿态和形状估计

Zhiying Li, Guanggang Geng, Yeying Jin, Shuyuan Lin, Fengyuan Ma, Zhaoxin Fan, Lili Wang

机构 * Beijing Advanced Innovation Center for Future Blockchain and Privacy Computing, School of Artificial Intelligence, Beihang University(北京未来区块链与隐私计算高级创新中心,人工智能学院,北航) College of Cyber Security, Jinan University(网络安全学院,暨南大学) National University of Singapore(新加坡国立大学) School of Computer Science and Information Engineering, Hefei University of Technology(计算机科学与信息工程学院,合肥工业大学) School of Computer Science and Engineering, Beihang University(计算机科学与工程学院,北航)

AI总结 本文提出LatentStealth方法,通过利用自然图像的结构化潜在表示,在潜在空间中生成并优化对抗扰动,实现无痕高效的对抗攻击,揭示当前系统关键安全漏洞。

Comments 10 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23742 2026-04-28 cs.SD

RTCFake: Speech Deepfake Detection in Real-Time Communication

RTCFake: 语音深度伪造检测中的实时通信

Jun Xue, Zhuolin Yi, Yihuan Huang, Yanzhen Ren, Yujie Chen, Cunhang Fan, Zicheng Su, Yonghong Zhang, Bo Cai

机构 * Key Laboratory of Aerospace Information Security and Trusted Computing, Ministry of Education(航空航天信息安全部与可信计算教育部重点实验室) School of Cyber Science and Engineering, Wuhan University(武汉大学计算机科学与工程学院) School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院) Beihang University(北京航空航天大学)

AI总结 本文提出RTCFake数据集,用于实时通信场景下的语音深度伪造检测,结合平台不变语义表征学习策略,提升跨平台泛化能力与噪声鲁棒性。

Comments Accepted by ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23551 2026-04-28 cs.CV

Spatiotemporal Degradation-Aware 3D Gaussian Splatting for Realistic Underwater Scene Reconstruction

时空降质感知的3D高斯点散布用于逼真水下场景重建

Shaohua Liu, Ning Gao, Zuoya Gu, Hongkun Dou, Yue Deng, Hongjue Li

机构 * School of Astronautics Beihang University Beijing China(航天学院 北航 北京 中国) School of Artificial Intelligence Beihang University Beijing China(人工智能学院 北航 北京 中国) Zhongguancun Academy Beijing China(中关村学院 北京 中国) Beihang University School of Astronautics Beijing China(北航 航天学院 北京 中国) Beihang University(北航) Zhongguancun Academy(中关村学院)

AI总结 本文提出MarineSTD-GS框架,通过建模时空降质实现逼真水下场景重建,引入内在高斯和降质高斯,结合时空降质建模模块,改进几何和外观估计,实验表明其在处理时空降质和合成新视角方面优于现有方法。

Comments 12 pages, 10 figures, 6 tables. Author version of the paper published in Proceedings of ACM Multimedia 2025

Journal ref Proceedings of the 33rd ACM International Conference on Multimedia (ACM MM 2025), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12971 2026-04-28 cs.RO

INHerit-SG: Incremental Hierarchical Semantic Scene Graphs with RAG-Style Retrieval

INHerit-SG:增量式层次语义场景图与RAG风格检索

YukTungSamuel Fang, Zhikang Shi, Jiabin Qiu, Zixuan Chen, Jieqi Shi, Hao Xu, Jing Huo, Yang Gao

机构 * State Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学) Beihang University(北京航空航天大学)

AI总结 本文提出INHerit-SG,通过异步双流架构构建RAG-ready知识库,解决复杂嵌入查询与持续语义图构建问题,实验表明在复杂查询中表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23314 2026-04-28 cs.CV

Learning from Noisy Prompts: Saliency-Guided Prompt Distillation for Robust Segmentation with SAM

从噪声提示中学习:基于显著性的提示蒸馏用于具有SAM的鲁棒分割

Jingxuan Kang, Ziqi Zhang, Shaoming Zheng, Shuang Li, Uday Bharat Patel, Alexander Harry Fitzhugh, Phillip Lung, Yusuf Kiberu, Nikesh Jathanna, Shahnaz Jamil-Copley, Bernhard Kainz, Chen Qin

机构 * Imperial College London(伦敦帝国学院) Beihang University(北航) National Health Service(国家卫生服务) University of Nottingham(诺丁汉大学)

AI总结 本文提出SPD框架,通过数据驱动的解剖先验知识和上下文提示蒸馏,提升在噪声提示下的分割鲁棒性,实验表明其在MRI和CT基准上优于现有方法。

Comments Accepted to CVPR 2026 (Findings Track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22828 2026-04-28 cs.CV cs.AI

MetaEarth3D: Unlocking World-scale 3D Generation with Spatially Scalable Generative Modeling

MetaEarth3D: 解锁世界尺度的3D生成与空间可扩展生成模型

Jinqi Cao, Zhiping Yu, Baihong Lin, Chenyang Liu, Zhenwei Shi, Zhengxia Zou

机构 * School of Astronautics, Beihang University(北航航天学院)

AI总结 MetaEarth3D通过空间可扩展生成模型实现世界尺度3D生成,解决传统生成模型在空间尺度上的局限,提升地球观测与模拟的超大空间智能能力。

Comments Project Page: https://jinqicao.github.io/metaearth3d/

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22777 2026-04-28 cs.AI cs.LG

An Intelligent Fault Diagnosis Method for General Aviation Aircraft Based on Multi-Fidelity Digital Twin and FMEA Knowledge Enhancement

基于多保真度数字孪生和FMEA知识增强的通用航空飞机智能故障诊断方法

Zhihuan Wei, Yang Hu, Xinhang Chen, Yiming Zhang, Jie Liu, Wei Wang

机构 * Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院) School of Mechanical Engineering, Zhejiang University(浙江大学机械工程学院) School of Reliability and Systems Engineering, Beihang University(北京航空航天大学可靠性与系统工程学院) Department of Mechanical Engineering, City University of Hong Kong(香港城市大学机械工程系)

AI总结 本文提出基于多保真度数字孪生的智能故障诊断框架,通过高保真飞行动态仿真、FMEA驱动故障注入、多保真残差特征提取和大语言模型增强的可解释报告生成模块,解决通用航空飞机故障诊断中的数据稀缺、故障类型多样和故障特征弱等问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23213 2026-04-28 cs.CL cs.AI

Scoring, Reasoning, and Selecting the Best! Ensembling Large Language Models via a Peer-Review Process

评分、推理与选择最佳!通过同行评审过程进行大型语言模型的集成

Zhijun Chen, Zeyu Ji, Qianren Mao, Hao Wu, Jinhuan Song, Junhang Cheng, Bangjie Qin, Zhuoran Li, Jingzheng Li, Kai Sun, Zizhe Wang, Yikun Ban, Zhu Sun, Xiangyang Ji, Hailong Sun

机构 * Beihang University, Beijing, China(北京航空航天大学) Zhongguancun Laboratory, Beijing, China(中关村实验室) Beijing University of Posts and Telecommunications(北京邮电大学) Hong Kong University of Science and Technology(香港科学与技术大学) Xi'an Jiaotong University, Xi'an, China(西安交通大学) Tsinghua University, Beijing, China(清华大学) Singapore University of Technology and Design(新加坡科技设计大学)

AI总结 本文提出LLM-PeerReview方法,通过同行评审机制集成多个大型语言模型,提升响应质量。实验显示其在多个数据集上优于Smoothie-Global,适用于事实性问答、数学推理等任务。

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.01579 2026-04-28 cs.CL cs.AI

AdaComp: Extractive Context Compression with Adaptive Predictor for Retrieval-Augmented Large Language Models

AdaComp: 基于自适应预测器的提取式上下文压缩用于检索增强型大语言模型

Qianchi Zhang, Hainan Zhang, Liang Pang, Hongwei Zheng, Zhiming Zheng

机构 * Beijing Advanced Innovation Center for Future Blockchain and Privacy Computing(北京未来区块链与隐私计算先进创新中心) School of Artificial Intelligence, Beihang University(北航人工智能学院) Beijing Academy of Blockchain and Edge Computing(北京区块链与边缘计算研究院) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)

AI总结 本文提出AdaComp,一种低开销的提取式上下文压缩方法,通过自适应确定压缩率来平衡RAG的效率与性能,实验表明其在保持性能的同时显著降低了推理成本。

Comments Accepted to KSEM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏