arXivDaily arXiv每日学术速递 周一至周五更新

作者

Cordelia Schmid

Computer Vision

共收录 226
2212.05922 2024-01-05 cs.CV cs.SD

Audiovisual Masked Autoencoders

Mariana-Iuliana Georgescu, Eduardo Fonseca, Radu Tudor Ionescu, Mario Lucic, Cordelia Schmid, Anurag Arnab

Comments ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.08506 2023-12-19 cs.CV cs.AI cs.LG

Does Visual Pretraining Help End-to-End Reasoning?

Chen Sun, Calvin Luo, Xingyi Zhou, Anurag Arnab, Cordelia Schmid

Comments NeurIPS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09237 2023-12-15 cs.CV

Pixel Aligned Language Models

Jiarui Xu, Xingyi Zhou, Shen Yan, Xiuye Gu, Anurag Arnab, Chen Sun, Xiaolong Wang, Cordelia Schmid

Comments Project page: https://jerryxu.net/PixelLLM

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.05582 2023-11-21 cs.CV

Image Matching with Scale Adjustment

Yves Dufournaud, Cordelia Schmid, Radu Horaud

Journal ref Computer Vision and Image Understanding, volume 93, 2004

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.08129 2023-11-03 cs.CV cs.AI cs.CL

AVIS: Autonomous Visual Information Seeking with Large Language Model Agent

Ziniu Hu, Ahmet Iscen, Chen Sun, Kai-Wei Chang, Yizhou Sun, David A Ross, Cordelia Schmid, Alireza Fathi

Comments Published on NeurIPS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.15596 2023-09-28 cs.RO cs.CV

PolarNet: 3D Point Clouds for Language-Guided Robotic Manipulation

Shizhe Chen, Ricardo Garcia, Cordelia Schmid, Ivan Laptev

Comments Accepted to CoRL 2023. Project website: https://www.di.ens.fr/willow/research/polarnet/

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.13952 2023-09-26 cs.CV cs.AI cs.CL cs.LG

VidChapters-7M: Video Chapters at Scale

Antoine Yang, Arsha Nagrani, Ivan Laptev, Josef Sivic, Cordelia Schmid

Comments Accepted at NeurIPS 2023 Track on Datasets and Benchmarks; Project Webpage: https://antoyang.github.io/vidchapters.html ; 31 pages; 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.14308 2023-08-30 cs.CV

WALDO: Future Video Synthesis using Object Layer Decomposition and Parametric Flow Prediction

Guillaume Le Moing, Jean Ponce, Cordelia Schmid

Comments Accepted to ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.12965 2023-08-25 cs.CV

POCO: 3D Pose and Shape Estimation with Confidence

Sai Kumar Dwivedi, Cordelia Schmid, Hongwei Yi, Michael J. Black, Dimitrios Tzionas

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.11062 2023-08-23 cs.CV cs.LG

UnLoc: A Unified Framework for Video Localization Tasks

Shen Yan, Xuehan Xiong, Arsha Nagrani, Anurag Arnab, Zhonghao Wang, Weina Ge, David Ross, Cordelia Schmid

Comments ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.13309 2023-08-21 cs.CV cs.AI

History Aware Multimodal Transformer for Vision-and-Language Navigation

Shizhe Chen, Pierre-Louis Guhur, Cordelia Schmid, Ivan Laptev

Comments Accepted in NeurIPS 2021; project page at https://cshizhe.github.io/projects/vln_hamt.html; corrected a typo

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.07282 2023-08-21 cs.CV cs.LG

Waffling around for Performance: Visual Classification with Random Words and Broad Concepts

Karsten Roth, Jae Myung Kim, A. Sophia Koepke, Oriol Vinyals, Cordelia Schmid, Zeynep Akata

Comments Accepted to ICCV 2023. Main paper with 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05602 2023-08-11 cs.CV cs.RO

Object Goal Navigation with Recursive Implicit Maps

Shizhe Chen, Thomas Chabal, Ivan Laptev, Cordelia Schmid

Comments Accepted to IROS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.15320 2023-07-31 cs.RO cs.AI cs.CV cs.LG

Robust Visual Sim-to-Real Transfer for Robotic Manipulation

Ricardo Garcia, Robin Strudel, Shizhe Chen, Etienne Arlaud, Ivan Laptev, Cordelia Schmid

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.11726 2023-06-21 cs.CV

How can objects help action recognition?

Xingyi Zhou, Anurag Arnab, Chen Sun, Cordelia Schmid

Comments CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.05392 2023-06-09 cs.CL

Modular Visual Question Answering via Code Generation

Sanjay Subramanian, Medhini Narasimhan, Kushal Khangaonkar, Kevin Yang, Arsha Nagrani, Cordelia Schmid, Andy Zeng, Trevor Darrell, Dan Klein

Comments ACL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.10140 2023-05-29 cs.CL cs.CV

Tackling Ambiguity with Images: Improved Multimodal Machine Translation and Contrastive Evaluation

Matthieu Futeral, Cordelia Schmid, Ivan Laptev, Benoît Sagot, Rachel Bawden

Comments Accepted to ACL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.06289 2023-05-11 cs.RO cs.CV cs.LG

Learning Video-Conditioned Policies for Unseen Manipulation Tasks

Elliot Chane-Sane, Cordelia Schmid, Ivan Laptev

Comments ICRA 2023. See the project webpage at https://www.di.ens.fr/willow/research/vip/

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.12160 2023-04-25 cs.CV

End-to-End Spatio-Temporal Action Localisation with Video Transformers

Alexey Gritsenko, Xuehan Xiong, Josip Djolonga, Mostafa Dehghani, Chen Sun, Mario Lučić, Cordelia Schmid, Anurag Arnab

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.11970 2023-04-25 cs.CV

gSDF: Geometry-Driven Signed Distance Functions for 3D Hand-Object Reconstruction

Zerui Chen, Shizhe Chen, Cordelia Schmid, Ivan Laptev

Comments Accepted by CVPR 2023. Project Page: https://zerchen.github.io/projects/gsdf.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.06708 2023-04-14 cs.CV cs.AI cs.CL

Verbs in Action: Improving verb understanding in video-language models

Liliane Momeni, Mathilde Caron, Arsha Nagrani, Andrew Zisserman, Cordelia Schmid

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.05173 2023-04-12 cs.CV cs.LG

Improving Image Recognition by Retrieving from Web-Scale Image-Text Data

Ahmet Iscen, Alireza Fathi, Cordelia Schmid

Comments Accepted to CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.03391 2023-04-10 cs.CV

Exposing and Mitigating Spurious Correlations for Cross-Modal Retrieval

Jae Myung Kim, A. Sophia Koepke, Cordelia Schmid, Zeynep Akata

Comments CVPR'23 MULA Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.01804 2023-04-05 cs.CV

Bridging the Gap between Model Explanations in Partially Annotated Multi-label Classification

Youngwook Kim, Jae Myung Kim, Jieun Jeong, Cordelia Schmid, Zeynep Akata, Jungwoo Lee

Comments CVPR2023 Camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.05221 2023-04-04 cs.CV cs.AI

REVEAL: Retrieval-Augmented Visual-Language Pre-Training with Multi-Source Multimodal Knowledge Memory

Ziniu Hu, Ahmet Iscen, Chen Sun, Zirui Wang, Kai-Wei Chang, Yizhou Sun, Cordelia Schmid, David A. Ross, Alireza Fathi

Comments Published on CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.16501 2023-03-30 cs.CV cs.SD eess.AS

AVFormer: Injecting Vision into Frozen Speech Models for Zero-Shot AV-ASR

Paul Hongsuck Seo, Arsha Nagrani, Cordelia Schmid

Comments CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.14115 2023-03-22 cs.CV cs.AI cs.CL cs.LG

Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning

Antoine Yang, Arsha Nagrani, Paul Hongsuck Seo, Antoine Miech, Jordi Pont-Tuset, Ivan Laptev, Josef Sivic, Cordelia Schmid

Comments CVPR 2023 Camera-Ready; Project Webpage: https://antoyang.github.io/vid2seq.html ; 18 pages; 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.02400 2023-03-17 cs.CV

Location-Aware Self-Supervised Transformers for Semantic Segmentation

Mathilde Caron, Neil Houlsby, Cordelia Schmid

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.09019 2023-03-08 cs.RO cs.AI cs.CV cs.LG

Learning Reward Functions for Robotic Manipulation by Observing Humans

Minttu Alakuijala, Gabriel Dulac-Arnold, Julien Mairal, Jean Ponce, Cordelia Schmid

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.09006 2023-02-17 cs.RO cs.LG

Enforcing the consensus between Trajectory Optimization and Policy Learning for precise robot control

Quentin Le Lidec, Wilson Jallet, Ivan Laptev, Cordelia Schmid, Justin Carpentier

详情

展开后加载摘要…

URL PDF HTML 收藏