arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

高校专区

University of Science and Technology of China(中国科学技术大学)

2026-03-18 至 2026-03-18 共收录 12
2603.16447 2026-03-18 cs.CV cs.GR

ProgressiveAvatars: Progressive Animatable 3D Gaussian Avatars

渐进式动画3D高斯化身:ProgressiveAvatars

Kaiwen Song, Jinkai Cui, Juyong Zhang

机构 * University of Science and Technology of China(中国科学技术大学)

AI总结 本文提出ProgressiveAvatars,基于自适应隐式细分的模板网格构建层次化3D高斯表示,实现网络和计算资源波动下的渐进式动画3D化身生成与渲染。

Comments Accepted to CVPR 2026, Project page: https://ustc3dv.github.io/ProgressiveAvatars/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16434 2026-03-18 cs.AI q-fin.TR

From Natural Language to Executable Option Strategies via Large Language Models

从自然语言到可执行期权策略的大型语言模型

Haochen Luo, Zhengzhao Lai, Junjie Xu, Yifan Li, Tang Pok Hin, Yuan Zhang, Chen Liu

机构 * City University of Hong Kong(香港城市大学) The Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳)) Shanghai University of Finance and Economics(上海财经大学) University of Science and Technology of China(中国科学技术大学)

AI总结 本文提出Option Query Language,通过语法规则将期权市场抽象为高层 primitives,使LLM能作为语义解析器生成可执行期权策略,并通过新数据集验证其优于直接生成方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16328 2026-03-18 cs.RO

ADAPT: Adaptive Dual-projection Architecture for Perceptive Traversal

ADAPT:面向感知穿越的自适应双投影架构

Shuo Shao, Tianchen Huang, Wei Gao, Shiwu Zhang

机构 * Department of Automation, University of Science and Technology of China(自动化系,中国科学技术大学) Institute of Humanoid Robots, Department of Precision Machinery and Precision Instrumentation, University of Science and Technology of China(人形机器人研究所,精密仪器与机械系,中国科学技术大学)

AI总结 ADAPT通过双投影地图提升复杂3D环境中的移动效率,通过可学习的动作范围实现感知范围的自适应调整,显著降低计算开销并提升训练速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16256 2026-03-18 cs.CV

When Thinking Hurts: Mitigating Visual Forgetting in Video Reasoning via Frame Repetition

当思考带来痛苦:通过帧重复缓解视频推理中的视觉遗忘

Xiaokun Sun, Yubo Wang, Haoyu Cao, Linli Xu

机构 * University of Science and Technology of China(中国科学技术大学) State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室)

AI总结 本文提出FrameRepeat框架,通过轻量级重复评分模块和Add-One-In策略,有效增强视频推理中的视觉线索,提升模型性能和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16219 2026-03-18 cs.CL

SpecSteer: Synergizing Local Context and Global Reasoning for Efficient Personalized Generation

SpecSteer:协同局部上下文与全局推理以实现高效个性化生成

Hang Lv, Sheng Liang, Hao Wang, Yongyue Zhang, Hongchao Gu, Wei Guo, Defu Lian, Yong Liu, Enhong Chen

机构 * University of Science and Technology of China(中国科学技术大学) Huawei Technologies Co., Ltd.(华为技术有限公司)

AI总结 本文提出SpecSteer框架,通过结合本地上下文与云端推理,解决个性化生成中的隐私与推理能力矛盾,实现高效生成并提升性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16210 2026-03-18 cs.AI

MOSAIC: Composable Safety Alignment with Modular Control Tokens

MOSAIC:通过模块化控制令牌实现可组合的安全对齐

Jingyu Peng, Hongyu Chen, Jiancheng Dong, Maolin Wang, Wenxi Li, Yuchen Li, Kai Zhang, Xiangyu Zhao

机构 * University of Science and Technology of China(中国科学技术大学) City University of Hong Kong(香港城市大学) Baidu Inc(百度公司) Minzu University of China(民族大学)

AI总结 MOSAIC通过可学习的控制令牌实现可组合的安全对齐,解决传统方法在动态安全规则下的局限性,实验显示其在降低过拒绝率的同时保持模型效用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16129 2026-03-18 cs.CV

Boosting Quantitive and Spatial Awareness for Zero-Shot Object Counting

提升零样本物体计数的定量与空间意识

Da Zhang, Bingyu Li, Feiyu Wang, Zhiyuan Zhao, Junyu Gao

机构 * Northwestern Polytechnical University(西北工业大学) Institute of Artificial Intelligence (TeleAI), China Telecom(中国电信人工智能研究院) University of Science and Technology of China(中国科学技术大学) Fudan University(复旦大学)

AI总结 本文提出QICA框架,通过引入协同提示策略和成本聚合解码器,提升零样本物体计数的定量感知和空间聚合能力,实验表明其在多个数据集上具有优越的泛化性能。

Comments Accepted to CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.14927 2026-03-18 cs.GR cs.LG

Masked BRep Autoencoder via Hierarchical Graph Transformer

基于分层图变换器的掩码BRep自编码器

Yifei Li, Kang Wu, Wenming Wu, Xiao-Ming Fu

机构 * University of Science and Technology of China(科学技术大学) Hefei University of Technology(合肥工业大学)

AI总结 本文提出一种自监督学习框架,通过自动学习CAD模型的表示以提升下游任务性能,采用掩码图自编码器和分层图变换器架构,实验表明模型在少量标注数据下表现优异。

Comments 27 pages, 11 figures. Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12633 2026-03-18 cs.CV cs.AI

DiG: Differential Grounding for Enhancing Fine-Grained Perception in Multimodal Large Language Model

DiG:通过差异 grounding 提升多模态大语言模型的细粒度感知

Zhou Tao, Shida Wang, Yongxiang Hua, Haoyu Cao, Linli Xu

机构 * University of Science and Technology of China(中国科学技术大学) State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室)

AI总结 本文提出DiG框架,通过学习相似图像对的差异识别提升多模态大语言模型的细粒度感知能力,实验表明其在多个视觉感知基准上表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07923 2026-03-18 cs.CV cs.AI

Exploring the Underwater World Segmentation without Extra Training

探索无需额外训练的水下世界分割

Bingyu Li, Tao Huo, Da Zhang, Zhiyuan Zhao, Junyu Gao, Xuelong Li

机构 * Institute of Artificial Intelligence (TeleAI), China Telecom, China(人工智能研究院(TeleAI),中国电信,中国) University of Science and Technology of China, China(中国科学技术大学,中国) Northwestern Polytechnical University, China(西北工业大学,中国)

AI总结 本文提出AquaOV255水下分割数据集及Earth2Ocean框架,通过几何引导视觉掩码生成和类别-视觉语义对齐模块,在无需额外训练的情况下实现高效的水下开放词汇分割。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15398 2026-03-18 cs.CV cs.AI

MARIS: Marine Open-Vocabulary Instance Segmentation with Geometric Enhancement and Semantic Alignment

MARIS: 基于几何增强与语义对齐的海洋开放词汇实例分割

Bingyu Li, Feiyu Wang, Da Zhang, Zhiyuan Zhao, Junyu Gao, Xuelong Li

机构 * University of Science and Technology of China(中国科学技术大学) Institute of Artificial Intelligence (TeleAI)(人工智能研究院(TeleAI)) China Telecom(中国电信) Fudan University(复旦大学) Northwestern Polytechnical University(西北工业大学)

AI总结 本文提出MARIS,首个大规模细粒度水下开放词汇分割基准,通过几何先验增强模块和语义对齐注入机制提升水下场景下的实例分割性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09551 2026-03-18 cs.CV

Label-supervised surgical instrument segmentation using temporal equivariance and semantic continuity

基于时间等变性和语义连续性的标注监督手术器械分割

Qiyuan Wang, Yanzhe Liu, Shang Zhao, Rong Liu, S. Kevin Zhou

机构 * School of Biomedical Engineering, Division of Life Sciences and Medicine, University of Science and Technology of China(中国科学技术大学生物医学工程学院,生命科学与医学系) Center for Medical Imaging, Robotics, Analytic Computing & Learning(MIRACLE), Suzhou Institute for Advanced Research, University of Science and Technology of China(中国科学技术大学苏州先进研究院医学影像中心、机器人与分析计算与学习中心) Key Laboratory of Precision and Intelligent Chemistry, University of Science and Technology of China(中国科学技术大学精密与智能化学重点实验室) Key Lab of Intelligent Information Processing of Chinese Academy of Sciences(CAS), Institute of Computing Technology, CAS(中国科学院智能信息处理重点实验室,计算技术研究所) Faculty of Hepato-Biliary-Pancreatic Surgery, The First Medical Center, Chinese PLA General Hospital(中国人民解放军总医院肝胆胰外科系)

AI总结 本文提出一种两阶段弱监督分割方法,通过时间等变性和语义连续性提升手术视频中器械分割的准确性,实验验证了其在两个手术数据集上的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏