arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7552 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7552 篇

2409.06305 2024-09-11 cs.CV 78%

High-Performance Few-Shot Segmentation with Foundation Models: An Empirical Study

Shijie Chang, Lihe Zhang, Huchuan Lu

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

Comments under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05904 2024-09-05 cs.CV 78%

Enhancing Representation in Radiography-Reports Foundation Model: A Granular Alignment Algorithm Using Masked Contrastive Learning

Weijian Huang, Cheng Li, Hong-Yu Zhou, Hao Yang, Jiarun Liu, Yong Liang, Hairong Zheng, Shaoting Zhang, Shanshan Wang

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

Journal ref Nature Communications 15, 7620 (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.08542 2024-09-04 cs.CV 78%

AIGCs Confuse AI Too: Investigating and Explaining Synthetic Image-induced Hallucinations in Large Vision-Language Models

Yifei Gao, Jiaqi Wang, Zhiyu Lin, Jitao Sang

专题命中 知识编辑与模型理解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.15676 2024-08-29 cs.SD eess.AS 78%

VoxInstruct: Expressive Human Instruction-to-Speech Generation with Unified Multilingual Codec Language Modelling

Yixuan Zhou, Xiaoyu Qin, Zeyu Jin, Shuoyi Zhou, Shun Lei, Songtao Zhou, Zhiyong Wu, Jia Jia

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted by ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01867 2024-08-06 cs.RO 78%

TrustNavGPT: Modeling Uncertainty to Improve Trustworthiness of Audio-Guided LLM-Based Robot Navigation

Xingpeng Sun, Yiran Zhang, Xindi Tang, Amrit Singh Bedi, Aniket Bera

专题命中 知识编辑与模型理解 :LLM(title,abstract)

Journal ref IROS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.08265 2024-07-29 cs.CV 78%

Coordinate-Aware Thermal Infrared Tracking Via Natural Language Modeling

Miao Yan, Ping Zhang, Haofei Zhang, Ruqian Hao, Juanxiu Liu, Xiaoyang Wang, Lin Liu

专题命中 知识编辑与模型理解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15605 2024-07-23 cs.CV 78%

Probing Fine-Grained Action Understanding and Cross-View Generalization of Foundation Models

Thinesh Thiyakesan Ponbagavathi, Kunyu Peng, Alina Roitberg

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.11335 2024-07-19 cs.CV 78%

LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction

Penghui Du, Yu Wang, Yifan Sun, Luting Wang, Yue Liao, Gang Zhang, Errui Ding, Yan Wang, Jingdong Wang, Si Liu

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments ECCV2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09128 2024-07-18 cs.CV 78%

Tokenize Anything via Prompting

Ting Pan, Lulu Tang, Xinlong Wang, Shiguang Shan

专题命中 知识编辑与模型理解 :prompting(title,abstract)

Comments code, model, and demo: https://github.com/baaivision/tokenize-anything

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09768 2024-06-14 cs.CV 78%

Contrastive Pretraining for Visual Concept Explanations of Socioeconomic Outcomes

Ivica Obadic, Alex Levering, Lars Pennig, Dario Oliveira, Diego Marcos, Xiaoxiang Zhu

专题命中 知识编辑与模型理解 :pretraining(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16094 2024-06-04 cs.CV 78%

PLUG: Revisiting Amodal Segmentation with Foundation Model and Hierarchical Focus

Zhaochen Liu, Limeng Qiao, Xiangxiang Chu, Tingting Jiang

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16311 2024-05-28 cs.HC 78%

Understanding Stakeholders' Perceptions and Needs Across the LLM Supply Chain

Agathe Balayn, Lorenzo Corti, Fanny Rancourt, Fabio Casati, Ujwal Gadiraju

专题命中 知识编辑与模型理解 :LLM(title,abstract)

Comments Paper accepted at the HCXAI workshop, co-located with CHI'24

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.13951 2024-05-24 cs.CV 78%

Text Prompting for Multi-Concept Video Customization by Autoregressive Generation

Divya Kothandaraman, Kihyuk Sohn, Ruben Villegas, Paul Voigtlaender, Dinesh Manocha, Mohammad Babaeizadeh

专题命中 知识编辑与模型理解 :prompting(title,abstract)

Comments Paper accepted to AI4CC Workshop at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.03592 2024-05-24 cs.CL cs.AI cs.LG 78%

ReFT: Representation Finetuning for Language Models

Zhengxuan Wu, Aryaman Arora, Zheng Wang, Atticus Geiger, Dan Jurafsky, Christopher D. Manning, Christopher Potts

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL、cs.AI、cs.LG

Comments preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.11683 2024-04-19 cs.RO cs.CV 78%

Unifying Scene Representation and Hand-Eye Calibration with 3D Foundation Models

Weiming Zhi, Haozhan Tang, Tianyi Zhang, Matthew Johnson-Roberson

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10193 2024-04-17 cs.CV 78%

Consistency and Uncertainty: Identifying Unreliable Responses From Black-Box Vision-Language Models for Selective Visual Question Answering

Zaid Khan, Yun Fu

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.07329 2024-04-17 cs.CV 78%

The Bias of Harmful Label Associations in Vision-Language Models

Caner Hazirbas, Alicia Sun, Yonathan Efroni, Mark Ibrahim

专题命中 知识编辑与模型理解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12693 2024-03-20 cs.CV 78%

As Firm As Their Foundations: Can open-sourced foundation models be used to create adversarial examples for downstream tasks?

Anjun Hu, Jindong Gu, Francesco Pinto, Konstantinos Kamnitsas, Philip Torr

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.14339 2024-03-07 cs.CV 78%

Towards Concept-based Interpretability of Skin Lesion Diagnosis using Vision-Language Models

Cristiano Patrício, Luís F. Teixeira, João C. Neves

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted for publication in IEEE ISBI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.12181 2024-01-23 cs.LG cs.AI cs.CL 78%

Universal Neurons in GPT2 Language Models

Wes Gurnee, Theo Horsley, Zifan Carl Guo, Tara Rezaei Kheirkhah, Qinyi Sun, Will Hathaway, Neel Nanda, Dimitris Bertsimas

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.09587 2024-01-09 cs.SE 78%

On the Reliability and Explainability of Language Models for Program Generation

Yue Liu, Chakkrit Tantithamthavorn, Yonghui Liu, Li Li

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted by ACM Transactions on Software Engineering and Methodology (TOSEM)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15543 2023-11-28 cs.CV 78%

Beyond Pixels: Exploring Human-Readable SVG Generation for Simple Images with Vision Language Models

Tong Zhang, Haoyang Liu, Peiyan Zhang, Yuxuan Cheng, Haohan Wang

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments 10 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.12188 2023-08-03 cond-mat.mtrl-sci physics.comp-ph 78%

Toward Accurate Interpretable Predictions of Materials Properties within Transformer Language Models

Vadim Korolev, Pavel Protsenko

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments 17 pages, 5 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.03307 2023-04-10 cs.CV eess.IV 78%

Vita-CLIP: Video and text adaptive CLIP via Multimodal Prompting

Syed Talal Wasim, Muzammal Naseer, Salman Khan, Fahad Shahbaz Khan, Mubarak Shah

专题命中 知识编辑与模型理解 :prompting(title,abstract)

Comments Accepted at CVPR-2023. Codes/models available at https://github.com/TalalWasim/Vita-CLIP

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.00182 2023-03-28 cs.CV 78%

Bidirectional Cross-Modal Knowledge Exploration for Video Recognition with Pre-trained Vision-Language Models

Wenhao Wu, Xiaohan Wang, Haipeng Luo, Jingdong Wang, Yi Yang, Wanli Ouyang

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.11100 2023-01-27 cs.CV cs.CY cs.HC 78%

Vision-Language Models Performing Zero-Shot Tasks Exhibit Gender-based Disparities

Melissa Hall, Laura Gustafson, Aaron Adcock, Ishan Misra, Candace Ross

专题命中 知识编辑与模型理解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.10578 2022-12-13 cs.CV 78%

ABINet++: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Spotting

Shancheng Fang, Zhendong Mao, Hongtao Xie, Yuxin Wang, Chenggang Yan, Yongdong Zhang

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted by TPAMI. Code is available at https://github.com/FangShancheng/ABINet-PP. arXiv admin note: substantial text overlap with arXiv:2103.06495 (conference version)

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.13407 2022-11-15 cs.CL cs.AI cs.LG 78%

CausaLM: Causal Model Explanation Through Counterfactual Language Models

Amir Feder, Nadav Oved, Uri Shalit, Roi Reichart

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL、cs.AI、cs.LG

Comments Our code and data are available at: https://amirfeder.github.io/CausaLM/ Accepted for publication in Computational Linguistics journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.14940 2022-03-29 cs.CV 78%

Learning to Prompt for Open-Vocabulary Object Detection with Vision-Language Model

Yu Du, Fangyun Wei, Zihe Zhang, Miaojing Shi, Yue Gao, Guoqi Li

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted by CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.06495 2021-03-12 cs.CV 78%

Read Like Humans: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Recognition

Shancheng Fang, Hongtao Xie, Yuxin Wang, Zhendong Mao, Yongdong Zhang

专题命中 知识编辑与模型理解 :language model(title,abstract)

Comments Accepted by CVPR 2021

详情

展开后加载摘要…

URL PDF HTML 收藏