arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

高校专区

New York University(纽约大学)

2026-04-10 至 2026-04-10 共收录 7
2604.08366 2026-04-10 cs.LG cs.AI cs.CV

Scaling-Aware Data Selection for End-to-End Autonomous Driving Systems

面向端到端自动驾驶系统的缩放感知数据选择

Tolga Dimlioglu, Nadine Chang, Maying Shen, Rafid Mahmood, Jose M. Alvarez

机构 * New York University(纽约大学) NVIDIA(英伟达) University of Ottawa(渥太华大学)

AI总结 本文提出MOSAIC框架,通过分域、拟合神经缩放规律和迭代优化数据混合,提升自动驾驶模型在EPDMS指标上的性能,比基线模型减少80%的数据使用。

Comments Accepted to CVPR 2026, 8 pages of main body and 10 pages of appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05674 2026-04-10 cs.CR cs.AI

Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark

迈向高效的攻击性安全LLM代理:超参数调优、LLM作为裁判以及一个轻量级CTF基准

Minghao Shao, Nanda Rani, Kimberly Milner, Haoran Xi, Meet Udeshi, Saksham Aggarwal, Venkata Sai Charan Putrevu, Sandeep Kumar Shukla, Prashanth Krishnamurthy, Farshad Khorrami, Ramesh Karri, Muhammad Shafique

机构 * New York University(纽约大学) New York University Abu Dhabi(纽约大学阿布扎比分校) Indian Institute of Technology Kanpur(印度理工学院坎普尔分校) International Institute of Information Technology Hyderabad(国际信息技术学院海得拉巴分校)

AI总结 本文系统研究了驱动代理成功的关键因素,提出CTFJudge框架和CTF Competency Index指标,分析超参数对性能的影响,并开放CTFTiny基准供研究。

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.00084 2026-04-10 cs.CL cs.AI

Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models

视觉-语言与大语言模型在胃肠病学中的表现:GPT、Claude、Llama、Phi、Mistral、Gemma及量化模型

Seyed Amir Ahmad Safavi-Naini, Shuhaib Ali, Omer Shahab, Zahra Shahhoseini, Thomas Savage, Sara Rafiee, Jamil S Samaan, Reem Al Shabeeb, Farah Ladak, Jamie O Yang, Juan Echavarria, Sumbal Babar, Aasma Shaukat, Samuel Margolis, Nicholas P Tatonetti, Girish Nadkarni, Bara El Kurdi, Ali Soroush

机构 * Icahn School of Medicine at Mount Sinai(西奈山伊坎医学院) University of Texas Health(德克萨斯大学健康科学中心) Virginia Hospital Center(弗吉尼亚医院中心) Shahid Beheshti University of Medical Sciences(沙希德·贝赫什提医科大学) Stanford University(斯坦福大学) Cedars-Sinai Medical Center(西达赛奈医疗中心) Inova Fairfax Medical Campus(伊诺瓦费尔法克斯医疗中心) University of California–Los Angeles(加州大学洛杉矶分校) NYU Grossman School of Medicine(纽约大学格罗斯曼医学院) Columbia University(哥伦比亚大学)

AI总结 本研究评估了大语言模型和视觉-语言模型在胃肠病学中的医学推理性能,比较了不同模型配置、参数及提示工程策略对性能的影响,发现专有模型在准确性上优于开源模型,且图像描述对视觉-语言模型性能有显著影响。

Comments Manuscript Pages: 34, Figures: 7, Tables: 2, Supplementary File Pages: 35, Data Transparency Statement: Code is available at: https://github.com/Sdamirsa/LLM-VLM-in-Gastroenterology . Study data from American College of Gastroenterology (ACG) are restricted and available upon request with ACG permission. Correction: updated abstract considering Llama3.1 results

Journal ref npj Digital Medicine 8, 797 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.05351 2026-04-10 cs.RO cs.CV

AnyImageNav: Any-View Geometry for Precise Last-Meter Image-Goal Navigation

AnyImageNav: 任意视角几何用于精确最后一步图像目标导航

Yijie Deng, Shuaihang Yuan, Yi Fang

机构 * NYUAD Center for Artificial Intelligence and Robotics (CAIR)(纽约大学阿布扎比分校人工智能与机器人中心) New York University Abu Dhabi, Electrical Engineering(纽约大学阿布扎比分校电气工程系) Embodied AI and Robotics (AIR) Lab, NYU Abu Dhabi(纽约大学阿布扎比分校具身人工智能与机器人实验室)

AI总结 本文提出AnyImageNav,通过语义到几何的级联方法实现精确6自由度相机姿态恢复,提升导航成功率和定位精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02670 2026-04-10 cs.CL

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting

攻破我:通过词性插入提示实现对对齐大语言模型的自我越狱

Devang Kulshreshtha, Hang Su, Haibo Jin, Chinmay Hegde, Haohan Wang

机构 * Amazon(亚马逊) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) New York University(纽约大学)

AI总结 本文提出自我越狱方法,通过词性插入提示对齐大语言模型进行攻击,实验表明SLIP在多个模型上实现高攻击成功率,提出SDM防御机制但发现其对自适应攻击策略仍不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23199 2026-04-10 stat.ML cs.LG

Rate-optimal Design for Anytime Best Arm Identification

任意时间最佳臂识别的最优设计

Junpei Komiyama, Kyoungseok Jang, Junya Honda

机构 * Stern School of Business, New York University(纽约大学斯特恩商学院) Department of AI, Chung-Ang University(中央大学人工智能系) Graduate School of Informatics, Kyoto University(京都大学信息学研究科) Machine Learning Department, Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学机器学习系) RIKEN AIP(理研革新智能研究中心)

AI总结 本文提出Almost Tracking算法,基于最小最大最优框架,无需提前预算或丢弃样本,在合成和现实数据中优于现有任意时间与固定预算算法。

Comments To appear in AISTATS2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20759 2026-04-10 cs.SD

Controllable Embedding Transformation for Mood-Guided Music Retrieval

可控嵌入变换用于情绪引导的音乐检索

Julia Wilkins, Jaehun Kim, Matthew E. P. Davies, Juan Pablo Bello, Matthew C. McCallum

机构 * SiriusXM-Pandora, USA(SiriusXM-Pandora,美国) New York University, New York, USA(纽约大学,纽约,美国)

AI总结 本文提出一种可控音乐嵌入变换框架,通过情绪标签引导种子音频嵌入到目标嵌入,保留其他音乐属性,实验表明在保持流派和乐器方面优于无训练基线。

Comments Preprint; under review

详情

展开后加载摘要…

URL PDF HTML 收藏