arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Fudan University(复旦大学)

2026-06-26 至 2026-06-26 共收录 11
2606.27339 2026-06-26 cs.CV 新提交

SAM2Matting: Generalized Image and Video Matting

SAM2Matting:通用图像与视频抠图

Ruiqi Shen, Guangquan Jie, Chang Liu, Henghui Ding

机构 * Fudan University(复旦大学) Shanghai University of Finance and Economics(上海财经大学)

AI总结 提出SAM2Matting框架,将VOS追踪器增强为高保真视频抠图,通过区域提议桥接和专用抠图头解耦任务,仅用图像训练即实现视频抠图新SOTA,支持多种提示类型并保持强时间一致性。

Comments ECCV 2026. Extended version. Project Page: https://henghuiding.com/SAM2Matting/

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26994 2026-06-26 cs.CV cs.AI 新提交

Event-Aware Instructed Assistant for Referring Video Segmentation

事件感知指令助手用于指代视频分割

Jinyu Liu, Henghui Ding, Shuting He, Yu-Gang Jiang

机构 * Institute of Big Data, College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与技术学院大数据研究所) Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身智能研究所) Shanghai University of Finance and Economics(上海财经大学)

AI总结 提出EVIS模型,通过可学习事件查询将视频分解为简单事件,结合对象-像素混合学习,实现层次化视频理解,在5个基准上表现优异。

Comments IEEE Transactions on Image Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26894 2026-06-26 cs.CV 新提交

Modeling Local, Global, and Cross-Modal Context in Multimodal 3D MRI

多模态3D MRI中的局部、全局和跨模态上下文建模

Minh Duc Do, Tillmann Rheude, Noel Kronenberg, Roland Eils, Benjamin Wild

机构 * Berlin Institute of Health at Charité - Universitätsmedizin Berlin(柏林健康研究所,柏林夏里特医学院) Health Data Science Unit, Heidelberg University Hospital and BioQuant(海德堡大学医院与BioQuant健康数据科学部) Intelligent Medicine Institute, Fudan University(复旦大学智能医学研究所) Department of Mathematics and Computer Science, Freie Universität Berlin(柏林自由大学数学与计算机科学系)

AI总结 提出MICViT,一种3D视觉Transformer,通过四种注意力机制显式建模模态内和跨模态的局部与全局交互,在多数据集脑龄预测任务中优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26850 2026-06-26 cs.GR cs.CV 新提交

Appearance-Preserving Refinement of Generated 3D Assets for Monochromatic Fabrication

面向单色制造的三维生成资产的外观保持优化

Chentao Shen, Chen Jia, Mingjie Huang, Zhuang Zhang, Haisen Zhao, Xiangru Huang

机构 * Zhejiang University(浙江大学) Westlake University(西湖大学) Fudan University(复旦大学) Hangzhou Dianzi University(杭州电子科技大学) Shandong University(山东大学)

AI总结 提出GenMF框架,通过外观导向的几何优化将纹理依赖的视觉线索转为几何着色效果,并引入可微应力正则化,在单色制造中保持外观细节并降低应力集中。

Comments under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26849 2026-06-26 cs.CV 新提交

Liquid Fusion of Heterogeneous Representations Towards General Salient Object Detection

异构表示的液态融合用于通用显著目标检测

Ke Chen, Ling Zhou, Guangqi Jiang, Gengshen Wu, Yi Liu, Shoukun Xu

机构 * School of Computer Science and Artificial Intelligence, Changzhou University(常州大学计算机科学与人工智能学院) College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院) Faculty of Data Science, City University of Macau(澳门城市大学数据科学学院)

AI总结 提出液态融合网络(LFNet),通过动态门控机制融合SSM(VMamba)和CNN(ConvNeXt)的异构表示,并设计显著引导上采样(SGU)算子,在五个任务上实现SOTA性能。

Comments 20 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26631 2026-06-26 cs.CV 新提交

Position Rebinding Cache Reuse: Replay-Free Visual Revisiting for Interleaved Multimodal Reasoning

位置重绑定缓存复用:交错多模态推理中无重放的视觉重访

Mengzhao Wang, Yanli Ji, Wangmeng Zuo, Peng Ye, Chongjun Tu

机构 * Sun Yat-sen University(中山大学) Shenzhen Loop Area Institute(深圳河套学院) Harbin Institute of Technology (HIT)(哈尔滨工业大学) Fudan University(复旦大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) The Chinese University of Hong Kong(香港中文大学)

AI总结 提出位置重绑定缓存复用(PRCR)框架,通过重绑历史视觉KV缓存的位置坐标,实现无重放的高效视觉重访,在多个多模态推理基准上平均准确率提升5%,计算量减少数万倍。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26549 2026-06-26 cs.AI cs.LG 新提交

PMDformer: Patch-Mean Decoupling Information Transformer for Long-term Forecasting

PMDformer: 用于长期预测的补丁均值解耦信息Transformer

Ao Hu, Liangjian Wen, Jiang Duan, Yong Dai, He Yan, Dongkai Wang, Jun Wang, Yukun Zhang, Ruoxi Jiang, Zenglin Xu

机构 * Southwestern University of Finance and Economics(西南财经大学) Shanghai Academy of AI for Science(上海人工智能研究院) Fudan University(复旦大学) Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) X-Humanoid Research Institute(X-Humanoid 研究院) Chengdu Everimaging Science and Technology Co., Ltd.(成都永新科技有限公司) Artificial Intelligence and Digital Finance Key Laboratory of Sichuan Province(四川省人工智能与数字金融重点实验室)

AI总结 提出补丁均值解耦方法分离趋势与残差形状,结合趋势恢复注意力和邻近变量注意力,提升长期时间序列预测的稳定性和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.17420 2026-06-26 cs.LG cs.AI cs.SI 版本更新

TransXion: A High-Fidelity Graph Benchmark for Realistic Anti-Money Laundering

TransXion: 一种高保真的图基准用于现实中的反洗钱

Keyang Chen, Mingxuan Jiang, Yongsheng Zhao, Zeping Li, Zaiyuan Chen, Weiqi Luo, Zhixin Li, Sen Liu, Yinan Jing, Guangnan Ye, Xihong Wu, Hongfeng Chai

机构 * College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院) Institute of Financial Technology, Fudan University(复旦大学金融技术研究所) Peking University(北京大学)

AI总结 TransXion通过整合正常活动的profile-aware模拟与非模板生成的非法子图,解决现有交易图数据集的稀疏性和模板偏差问题,提供更真实的反洗钱检测测试环境。

Journal ref Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining V.2 (KDD '26), 8707-8718, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16220 2026-06-26 cs.LG 版本更新

SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series Forecasting

SEMixer: 语义增强的MLP-Mixer用于多尺度混合和长期时间序列预测

Xu Zhang, Qitong Wang, Peng Wang, Wei Wang

机构 * Shanghai Key Laboratory of Data Science, College of Computer Science and Artificial Intelligence Fudan University(上海数据科学 key 实验室,复旦大学计算机科学与人工智能学院) Harvard University(哈佛大学)

AI总结 提出SEMixer模型,通过随机注意力机制和多尺度渐进混合链,有效建模多尺度时间依赖并解决语义鸿沟问题,在10个公开数据集和真实无线网络数据上取得优异性能。

Comments This work is accepted by the proceedings of the ACM Web Conference 2026 (WWW 2026). The code is available at the link https://github.com/Meteor-Stars/SEMixer

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08275 2026-06-26 q-bio.NC cs.CL 版本更新

Linguistics and Human Brain: A Perspective of Computational Neuroscience

语言学与人脑:计算神经科学的视角

Fudong Zhang, Bo Chai, Yujie Wu, Wai Ting Siok, Nizhuan Wang

机构 * Institute of AI and Robotics, College of Intelligent Robotics and Advanced Manufacturing, Fudan University(人工智能与机器人研究院,智能机器人与先进制造学院,复旦大学) Department of Language Science and Technology, The Hong Kong Polytechnic University(语言科学与技术系,香港理工大学) Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学)

AI总结 本文从计算神经科学视角,通过建模、模拟和数据分析将语言层次动态结构转化为可测试神经模型,并利用深度学习和大语言模型探索语言处理的神经基础。

Journal ref Cognitive Neurodynamics, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09960 2026-06-26 cs.LG cs.AI 版本更新

Limited Reference, Reliable Generation: A Two-Component Framework for Tabular Data Generation in Low-Data Regimes

有限参考,可靠生成:低数据场景下的表格数据生成双组件框架

Mingxuan Jiang, Keyang Chen, Yongxin Wang, Yongsheng Zhao, Ziyue Dai, Yicun Liu, Zeping Li, Qiuyang Zhang, Hongyi Nie, Hongbin Zhu, Sen Liu, Guangnan Ye, Hongfeng Chai

机构 * School of Computer Science, Fudan University(复旦大学计算机科学学院) Institute of Financial Technology, Fudan University(复旦大学金融技术研究院) Northwestern Polytechnical University(西北工业大学)

AI总结 提出ReFine框架,通过从可解释模型提取符号规则嵌入提示引导生成,并采用双粒度过滤减少局部冗余,在低数据场景下实现稳健的表格数据生成,平均提升7.48%。

详情

展开后加载摘要…

URL PDF HTML 收藏