arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

高校专区

University of Science and Technology of China(中国科学技术大学)

2026-03-31 至 2026-03-31 共收录 12
2603.28182 2026-03-31 cs.CV

A Closer Look at Cross-Domain Few-Shot Object Detection: Fine-Tuning Matters and Parallel Decoder Helps

对跨域少样本目标检测的深入研究:微调至关重要,平行解码器有所帮助

Xuanlong Yu, Youyang Sha, Longfei Liu, Xi Shen, Di Yang

机构 * Intellindust AI Lab(Intellindust AI实验室) Suzhou Institute for Advanced Research, USTC(中国科学技术大学苏州高等研究院)

AI总结 本文提出混合集成解码器提升微调泛化能力,结合统一渐进微调框架和plateau-aware学习率调度,验证了在CD-FSOD、ODinW-13和RF100-VL数据集上的有效性,尤其在10-shot设置中优于SAM3。

Comments Accepted at CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.28162 2026-03-31 cs.CV

ColorFLUX: A Structure-Color Decoupling Framework for Old Photo Colorization

ColorFLUX:一种用于旧照片着色的结构-颜色解耦框架

Bingchen Li, Zhixin Wang, Fan Li, Jiaqi Xu, Jiaming Guo, Renjing Pei, Xin Li, Zhibo Chen

机构 * University of Science and Technology of China(中国科学技术大学) Huawei Noah’s Ark Lab(华为诺亚方舟实验室)

AI总结 本文提出基于生成扩散模型FLUX的新型旧照片着色框架,通过结构-颜色解耦策略和改进的直接偏好优化策略,提升旧照片着色的准确性与质量。

Comments Accepted by CVPR26

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.28046 2026-03-31 cs.AI

Dogfight Search: A Swarm-Based Optimization Algorithm for Complex Engineering Optimization and Mountainous Terrain Path Planning

狗群搜索:一种基于群体的优化算法,用于复杂工程优化和山地地形路径规划

Yujing Sun, Jie Cai, Xingguo Xu, Yuansheng Gao, Lei Zhang, Kaichen Ouyang, Zhanyu Liu

机构 * College of Science, Liaoning Technical University(辽宁工程技术大学理学院) School of Mathematical Sciences, Dalian University of Technology(大连理工大学数学科学学院) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) Department of Physics, University of Science and Technology of China(中国科学技术大学物理系)

AI总结 本文提出了一种名为狗群搜索(DoS)的新型元启发式算法,通过动力学位移积分方程构建搜索机制,在复杂工程优化和山地路径规划中表现出色,优于7种先进算法。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21643 2026-03-31 cs.CV

Omni-Weather: A Unified Multimodal Model for Weather Radar Understanding and Generation

Omni-Weather: 一种统一的多模态模型用于天气雷达理解和生成

Zhiwang Zhou, Yuandong Pu, Xuming He, Yidi Liu, Yixin Chen, Junchao Gong, Xiang Zhuang, Wanghan Xu, Qinglong Cao, Shixiang Tang, Yihao Liu, Wenlong Zhang, Lei Bai

机构 * Tongji University(同济大学) Shanghai AI Laboratory(上海人工智能实验室) Shanghai Jiao Tong University(上海交通大学) Zhejiang University(浙江大学) University of Science and Technology of China(中国科学技术大学) UCLA(加州大学洛杉矶分校)

AI总结 Omni-Weather通过统一架构整合天气生成与理解,采用共享自注意力机制和因果推理数据集提升生成质量与可解释性,实验显示其在天气生成和理解上达到最新水平。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.27625 2026-03-31 cs.CV

Clore: Interactive Pathology Image Segmentation with Click-based Local Refinement

Clore:基于点击的病理图像分割交互式局部细化

Tiantong Wang, Minfan Zhao, Jun Shi, Hannan Wang, Yue Dai

机构 * School of Computer Science and Technology, University of Science and Technology of China(中国科学技术大学计算机科学与技术学院) School of Artificial Intelligence and Data Science, University of Science and Technology of China(中国科学技术大学人工智能与数据科学学院)

AI总结 本文提出Clore方法,通过分层交互范式提升病理图像分割精度,减少交互次数,实现高效准确的局部细化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.27301 2026-03-31 cs.CV

Dual-Path Learning based on Frequency Structural Decoupling and Regional-Aware Fusion for Low-Light Image Super-Resolution

基于频率结构解耦和区域感知融合的双路径学习用于低光照图像超分辨率

Ji-Xuan He, Jia-Cheng Zhao, Feng-Qi Cui, Jinyang Huang, Yang Liu, Sirui Zhao, Meng Li, Zhi Liu

机构 * Hefei University of Technology(合肥工业大学) University of Science and Technology of China(中国科学技术大学) Zhejiang University(浙江大学) The University of Electro-Communications(电气通信大学)

AI总结 本文提出DTP框架,通过频率感知解耦和区域感知融合提升低光照图像超分辨率,改进PSNR、SSIM和LPIPS指标。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.27165 2026-03-31 cs.CV

RiskProp: Collision-Anchored Self-Supervised Risk Propagation for Early Accident Anticipation

RiskProp: 基于碰撞锚定的自监督风险传播用于早期事故预警

Yiyang Zou, Tianhao Zhao, Peilun Xiao, Hongyu Jin, Longyu Qi, Yuxuan Li, Liyin Liang, Yifeng Qian, Chunbo Lai, Yutian Lin, Zhihui Li, Yu Wu

机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) Zhongguancun Academy(中关村学院) Didi Chuxing(滴滴出行) University of Science and Technology of China(中国科学技术大学)

AI总结 RiskProp通过自监督学习消除对异常起始帧的依赖,利用可靠标注的碰撞帧建模时间风险演变,提升早期事故预警的准确性和可解释性。

Comments Accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08021 2026-03-31 cs.RO cs.CV

AffordGrasp: Cross-Modal Diffusion for Affordance-Aware Grasp Synthesis

AffordGrasp:跨模态扩散用于感知意识抓取合成

Xiaofei Wu, Yi Zhang, Yumeng Liu, Yuexin Ma, Yujiao Shi, Xuming He

机构 * ShanghaiTech University(上海科技大学) Shanghai Engineering Research Center of Intelligent Vision and Imaging(上海智能视觉与成像工程技术研究中心) University of Science and Technology of China(中国科学技术大学)

AI总结 本文提出AffordGrasp,通过跨模态扩散模型生成准确反映物体几何和用户指令的抓取姿态,提升AR/VR和具身AI中的手-物交互质量。

Comments CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.01331 2026-03-31 cs.CY cs.CL cs.LG

AppellateGen: A Benchmark for Appellate Legal Judgment Generation

AppellateGen:上诉法律判决生成的基准测试

Hongkun Yang, Lionel Z. Wang, Wei Fan, Yiran Hu, Lixu Wang, Chenyu Liu, Yu Zeng, Shenghong Fu, Lei Gong, Zhengxin Zhang, Haoyang Li, Jiexin Zheng, Xin Xu

机构 * Nanyang Technological University(南洋理工大学) Ocean University of China(中国海洋大学) The Hong Kong Polytechnic University(香港理工大学) Hong Kong University of Science and Technology(香港科技大学) Tsinghua University(清华大学) University of Science and Technology of China(中国科学技术大学) Cornell University(康奈尔大学)

AI总结 本文提出AppellateGen基准测试,包含7351个案例对,用于生成上诉阶段的法律判决,通过模拟司法流程验证模型在上诉推理中的逻辑一致性与挑战。

Comments 15 pages, 4 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05422 2026-03-31 cs.CV

ParaUni: Enhance Generation in Unified Multimodal Model with Reinforcement-driven Hierarchical Parallel Information Interaction

ParaUni: 通过强化驱动的分层并行信息交互增强统一多模态模型的生成

Jiangtong Tan, Lin Liu, Jie Huanng, Xiaopeng Zhang, Qi Tian, Feng Zhao

机构 * MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(中国科学技术大学,教育部类脑智能感知与认知重点实验室) Huawei Inc.(华为技术有限公司)

AI总结 ParaUni通过并行提取多层视觉语言模型特征并结合强化学习动态调整机制,提升统一多模态模型的生成能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13688 2026-03-31 cs.GR cs.AI

CraftMesh: High-Fidelity Generative Mesh Manipulation via Poisson Seamless Fusion

CraftMesh: 通过泊松无缝融合实现高保真的生成网格操控

James Jincheng, Yuxiao Wu, Youcheng Cai, Ligang Liu

机构 * Hefei Thomas School(合肥托马斯学校) University of Science and Technology of China(中国科学技术大学)

AI总结 CraftMesh通过将网格编辑分解为2D和3D生成模型的联合流程,结合泊松几何融合与纹理和谐技术,实现复杂几何的高保真生成与无缝融合,优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.26800 2026-03-31 cs.LG cs.AI physics.flu-dyn

DSO: Dual-Scale Neural Operators for Stable Long-term Fluid Dynamics Forecasting

DSO:双尺度神经算子用于稳定长期流体动力学预测

Huanshuo Dong, Hao Wu, Hong Wang, Qin-Yi Zhang, Zhezheng Hao

机构 * Institute for Clarity in Documentation(文档清晰度研究所) Inria Paris-Rocquencourt(法国国家信息与自动化研究所巴黎-罗康库尔) Rajiv Gandhi University(拉吉夫·甘地大学) Tsinghua University(清华大学) Palmer Research Laboratories(帕尔默研究实验室) University of Science and Technology of China(中国科学技术大学) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Zhejiang University(浙江大学)

AI总结 本文提出DSO双尺度神经算子,通过分离局部细节和全局趋势处理,提升长期流体预测的稳定性和精度。

详情

展开后加载摘要…

URL PDF HTML 收藏