KANGURA: Kolmogorov-Arnold Network-Based Geometry-Aware Learning with Unified Representation Attention for 3D Modeling of Complex Structures
Mohammad Reza Shafie, Morteza Hajiabadi, Hamed Khosravi, Mobina Noori, Imtiaz Ahmed
机构
*
Department of Electrical Engineering, Iran University of Science and Technology(伊朗科学技术大学电气工程系)
;
Department of Computer Engineering, Iran University of Science and Technology(伊朗科学技术大学计算机工程系)
;
Department of Industrial & Management Systems Engineering, West Virginia University(西弗吉尼亚大学工业与管理系统工程系)
;
H. Milton Stewart School of Industrial and Systems Engineering, Georgia Institute of Technology(佐治亚理工学院H.米尔顿·斯图尔特工业与系统工程学院)
;
Department of Computer Science, University of California, Davis(加州大学戴维斯分校计算机科学系)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
HAOMO.AI Technology Co., Ltd.(HAOMO.AI技术有限公司)
;
National University of Defense Technology(国防科技大学)
;
Beijing Institute of Technology(北京理工大学)
Caption This, Reason That: VLMs Caught in the Middle
Zihan Weng, Lucas Gomez, Taylor Whittington Webb, Pouya Bashivan
机构
*
Integrated Program in Neuroscience (IPN) McGill University(神经科学联合计划 麦吉尔大学)
;
Mila, University of Montreal(蒙特利尔大学Mila)
;
Microsoft Research USA(微软研究院美国总部)
;
Department of Physiology McGill University(生理学系 麦吉尔大学)
TESGNN: Temporal Equivariant Scene Graph Neural Networks for Efficient and Robust Multi-View 3D Scene Understanding
Quang P. M. Pham, Khoi T. N. Nguyen, Lan C. Ngo, Truong Do, Dezhen Song, Truong-Son Hy
机构
*
Department of Robotics(机器人学系)
;
Mohamed bin Zayed University of Artificial Intelligence(莫扎德·本·扎伊德人工智能大学)
;
College of Engineering and Computer Science(工程与计算机科学学院)
;
VinUniversity(文大学)
;
Department of Computer Science(计算机科学系)
;
The University of Alabama at Birmingham(阿拉巴马大学伯明翰分校)
专题命中
空间理解
:point cloud(abstract);分类 cs.CV
CommentsarXiv admin note: text overlap with arXiv:2407.00609
Uncertainty-Informed Active Perception for Open Vocabulary Object Goal Navigation
Utkarsh Bajpai, Julius Rückin, Cyrill Stachniss, Marija Popović
机构
*
Center for Robotics, University of Bonn(波恩大学机器人中心)
;
MAVLab, TU Delft(代尔夫特理工大学MAVLab)
;
Lamarr Institute for Machine Learning and Artificial Intelligence(机器学习与人工智能拉马尔研究所)
ERUPT: An Open Toolkit for Interfacing with Robot Motion Planners in Extended Reality
Isaac Ngui, Courtney McBeth, André Santos, Grace He, Katherine J. Mimnaugh, James D. Motes, Luciano Soares, Marco Morales, Nancy M. Amato
机构
*
Parasol Lab, Siebel School of Computing and Data Science, University of Illinois Urbana-Champaign(Parasol实验室,Siebel计算与数据科学学院,伊利诺伊大学厄巴纳-香槟分校)
;
Insper
;
Department of Computer Science at Instituto Tecnológico Autónomo de México (ITAM)(墨西哥自主技术学院计算机科学系)
机构
*
McGill University(麦吉尔大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Shanghai Jiao Tong University(上海交通大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
University of Florida(佛罗里达大学)
;
The University of Hong Kong(香港大学)
OmniScene: Attention-Augmented Multimodal 4D Scene Understanding for Autonomous Driving
Pei Liu, Hongliang Lu, Haichao Liu, Haipeng Liu, Xin Liu, Ruoyu Yao, Shengbo Eben Li, Jun Ma
机构
*
The Hong Kong University of Science and Technology(香港科技大学)
;
Li Auto Inc.
;
the School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动系统学院)
Improving Robotic Manipulation with Efficient Geometry-Aware Vision Encoder
An Dinh Vuong, Minh Nhat Vu, Ian Reid
机构
*
Department of Computer Vision, Mohammed bin Zayed University of Artificial Intelligence(计算机视觉系,穆罕默德·本·扎耶德人工智能大学)
;
AIT Austrian Institute of Technology GmbH(奥地利技术研究所)
TopoNav: Topological Graphs as a Key Enabler for Advanced Object Navigation
Peiran Liu, Qiang Zhang, Daojie Peng, Lingfeng Zhang, Yihao Qin, Hang Zhou, Jun Ma, Renjing Xu, Yiding Ji
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Beijing Innovation Center of Humanoid Robotics Co., Ltd.(北京人形机器人创新中心有限公司)
;
Shenzhen International Graduate School, Tsinghua University(深圳国际研究生学院,清华大学)