arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

高校专区

Zhejiang University(浙江大学)

2026-03-26 至 2026-03-26 共收录 11
2603.24578 2026-03-26 cs.CV eess.IV

Vision-Language Models vs Human: Perceptual Image Quality Assessment

视觉-语言模型与人类:感知图像质量评估

Imran Mehmood, Imad Ali Shah, Ming Ronnier Luo, Brian Deegan

机构 * School of Engineering, University of Galway(Galway大学工程学院) State Key Laboratory of Extreme Photonics and Instrumentation, Zhejiang University(浙江大学极端光信息获取国家重点实验室)

AI总结 本文研究视觉语言模型是否能近似人类对图像质量的感知判断,通过对比心理物理数据,发现模型在颜色丰富度上表现优异,但在对比度上表现较差,且模型一致性与人类对齐性存在矛盾。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20732 2026-03-26 cs.LG cs.CV

Continual GUI Agents

连续GUI代理

Ziwei Liu, Borui Kang, Hangjie Yuan, Zixiang Zhao, Wei Li, Yifan Zhu, Tao Feng

机构 * Department of Computer Science and Technology, Tsinghua University, China(计算机科学与技术系,清华大学,中国) College of Computer Science, Zhejiang University, China(浙江大学计算机科学学院,中国) College of Computer Science, Beijing University of Posts and Telecommunications, China(北京邮电大学计算机科学学院,中国)

AI总结 本文提出连续GUI代理任务,通过引入GUI-Anchoring in Flux框架,解决GUI分布变化时持续学习稳定性问题,实验显示其优于现有方法。

Comments Code is available at: https://github.com/xavierliu34/GUI-AiF

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24051 2026-03-26 cs.CL

FinToolSyn: A forward synthesis Framework for Financial Tool-Use Dialogue Data with Dynamic Tool Retrieval

FinToolSyn: 一种用于金融工具使用对话数据的前向合成框架,具有动态工具检索

Caishuang Huang, Yang Qiao, Rongyu Zhang, Junjie Ye, Pu Lu, Wenxi Wu, Meng Zhou, Xiku Du, Tao Gui, Qi Zhang, Xuanjing Huang

机构 * College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院) FiT, Tencent(腾讯FiT) Zhejiang University(浙江大学)

AI总结 本文提出FinToolSyn框架,通过动态工具检索生成高质量金融对话数据,构建43,066个工具库并合成148k对话实例,提升金融场景下的工具调用能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23512 2026-03-26 cs.CL cs.AI cs.IR

S-Path-RAG: Semantic-Aware Shortest-Path Retrieval Augmented Generation for Multi-Hop Knowledge Graph Question Answering

S-Path-RAG:基于语义的最短路径检索增强生成用于多跳知识图谱问答

Rong Fu, Yemin Wang, Tianxiang Xu, Yongtai Liu, Weizhi Tang, Wangyu Wu, Xiaowen Ma, Simon Fong

机构 * University of Macau(澳门大学) Xiamen University(厦门大学) Peking University(北京大学) Hanyang University(翰阳大学) University of Liverpool(利物浦大学) Zhejiang University(浙江大学)

AI总结 S-Path-RAG通过混合加权k最短路径、束搜索和约束随机游走策略,实现语义感知的多跳知识图谱问答,提升检索效率与生成准确性。

Journal ref WWW 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23383 2026-03-26 cs.CV

From Feature Learning to Spectral Basis Learning: A Unifying and Flexible Framework for Efficient and Robust Shape Matching

从特征学习到谱基学习:一种统一且灵活的框架,用于高效且鲁棒的形状匹配

Feifan Luo, Hongyang Chen

机构 * College of Computer Science and Technology, Zhejiang University, China(浙江大学计算机科学与技术学院,中国) Research Center for Computational Earth and Space Science, Zhejiang Lab, China(浙江实验室计算地球与空间科学研究中心,中国)

AI总结 本文提出一种统一的框架,通过学习谱基来提升形状匹配的效率和鲁棒性,引入了新的热扩散模块和无监督损失函数,实现了特征提取与基函数的联合优化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01718 2026-03-26 cs.RO cs.CV

Unified Diffusion VLA: Vision-Language-Action Model via Joint Discrete Denoising Diffusion Process

统一扩散VLA:通过联合离散去噪扩散过程实现视觉-语言-动作模型

Jiayi Chen, Wenxuan Song, Pengxiang Ding, Ziyang Zhou, Han Zhao, Feilong Tang, Donglin Wang, Haoang Li

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Westlake University(西交大学) Zhejiang University(浙江大学) Monash University(墨尔本大学)

AI总结 本文提出统一扩散VLA模型,通过联合离散去噪扩散过程实现视觉语言动作的协同生成与执行,提升多模态任务的效率与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11306 2026-03-26 cs.RO

Rotor-Failure-Aware Quadrotors Flight in Unknown Environments

具有旋转故障意识的四旋翼在未知环境中的飞行

Xiaobin Zhou, Miao Wang, Chengao Li, Can Cui, Ruibin Zhang, Yongchao Wang, Chao Xu, Fei Gao

机构 * School of Robotics and Automation, Nanjing University(南京大学机器人与自动化学院) Institute of Cyber-Systems and Control, College of Control Science and Engineering, Zhejiang University(浙江大学控制系统研究所) Department of Aeronautical and Aviation Engineering, The Hong Kong Polytechnic University(香港理工大学航空工程系) School of Aeronautic Science and Engineering, Beihang University(北航航空科学与工程学院)

AI总结 本文提出一种具有旋转故障意识的四旋翼导航系统,通过复合故障检测与诊断非线性模型预测控制器和LiDAR平台,在未知复杂环境中实现旋转故障的快速检测与飞行稳定。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22460 2026-03-26 cs.AI

GeoSketch: A Neural-Symbolic Approach to Geometric Multimodal Reasoning with Auxiliary Line Construction and Affine Transformation

GeoSketch: 一种基于神经符号的方法,用于几何多模态推理中的辅助线构造与仿射变换

Shichao Weng, Zhiqiang Wang, Yuhua Zhou, Rui Lu, Ting Liu, Zhiyang Teng, Xiaozhang Liu, Hanmeng Liu

机构 * Fudan University(复旦大学) IFLYTEK CO.LTD(若lytek有限公司) Zhejiang University(浙江大学) The Hong Kong University of Science and Technology(香港科学与技术大学) National University of Defense Technology(国防科技大学) Hainan University(海南大学)

AI总结 GeoSketch通过整合感知、符号推理和绘图动作模块,实现动态几何推理,提升多模态推理的准确性和解决问题的成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03611 2026-03-26 cs.LG cs.DB

Learning-based Sketches for Frequency Estimation in Data Streams without Ground Truth

基于学习的 sketches 在无地面真实数据流中频率估计

Xinyu Yuan, Yan Qiao, Meng Li, Zhenchun Wei, Cuiying Feng, Zonghui Wang, Wenzhi Chen

机构 * School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) Key Laboratory of Knowledge Engineering with Big Data, Hefei University of Technology(合肥工业大学大数据知识工程重点实验室) School of Information and Software Engineering, University of Electronic Science and Technology of China(电子科技大学信息与软件工程学院)

AI总结 本文提出UCL-sketch,一种无需地面真实数据的基于学习的频率估计方法,通过在线训练和逻辑结构化估计桶实现高效准确的频率估计,实验显示其在精度和分布上优于现有方法,且在内存受限情况下接近理想 oracle 的性能。

Comments Accepted as a regular paper at IEEE TKDE

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.01111 2026-03-26 cs.LG cs.AI stat.ML

Proximity Matters: Local Proximity Enhanced Balancing for Treatment Effect Estimation

近邻重要性:局部近邻增强的平衡方法用于处理效应估计

Hao Wang, Zhichao Chen, Zhaoran Liu, Xu Chen, Haoxuan Li, Zhouchen Lin

机构 * College of Control Science and Engineering, Zhejiang University(浙江大学控制科学与工程学院) School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学院)

AI总结 本文提出CFR-Pro方法,通过引入基于最优传输的成对近邻正则化来增强HTE估计中的表示平衡,有效缓解处理选择偏差,并在实验中优于其他方法。

Comments Accepted as a poster in SIGKDD 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.11039 2026-03-26 cs.LG cs.AI stat.ML

Entire Space Counterfactual Learning for Reliable Content Recommendations

全空间反事实学习用于可靠的内容推荐

Hao Wang, Zhichao Chen, Zhaoran Liu, Haozhe Li, Degui Yang, Xinggao Liu, Haoxuan Li

机构 * State Key Laboratory of Industrial Control Technology, College of Control Science and Engineering, Zhejiang University(工业控制技术国家重点实验室,控制科学与工程学院,浙江大学) School of Automation, Central South University(自动化学院,中南大学) Center for Data Science, Peking University(数据科学中心,北京大学)

AI总结 本文提出全空间反事实多任务模型ESCM²,通过引入反事实风险最小化模块,解决推荐系统中点击转化率估计的内在偏差和虚假独立性问题,提升推荐性能。

Comments This submission is an extension of arXiv:2204.05125

详情

展开后加载摘要…

URL PDF HTML 收藏