An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation
多时相指代分割的开源基准与基线
Bingyu Li, Da Zhang, Tao Huo, Zhiyuan Zhao, Junyu Gao, Xuelong Li
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Institute of Artificial Intelligence (TeleAI)(人工智能研究所)
;
China Telecom(中国电信)
;
School of Artificial Intelligence, Optics and Electronics (iOPEN)(人工智能、光学与电子学院)
;
Northwestern Polytechnical University(西北工业大学)
机构
*
VA Bedford Health Care(VA贝德福德医疗中心)
;
UMass Amherst(马萨诸塞大学阿默斯特分校)
;
UMass Lowell(马萨诸塞大学洛厄尔分校)
;
Yale University(耶鲁大学)
;
National University of Singapore(新加坡国立大学)
;
Yale School of Medicine(耶鲁医学院)
Vision-language models for chest radiography do not always need the image
胸部X光片的视觉-语言模型并不总是需要图像
Mahshad Lotfinia, Sebastian Ziegelmayer, Lisa Adams, Daniel Truhn, Andreas Maier, Soroosh Tayebi Arasteh
机构
*
Pattern Recognition Lab, Friedrich-Alexander-Universität Erlangen-Nürnberg(弗里德里希-亚历山大-埃尔朗根-纽伦堡大学模式识别实验室)
;
Department of Diagnostic and Interventional Radiology, TUM University Clinic, School of Medicine and Health, Klinikum rechts der Isar, Technical University of Munich(慕尼黑工业大学医学院与健康学院伊萨尔河右岸医院诊断与介入放射学系)
;
Lab for AI in Medicine, RWTH Aachen University(亚琛工业大学医学人工智能实验室)
;
Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(亚琛工业大学医院诊断与介入放射学系)
机构
*
Department of Computer Science, Rutgers University-New Brunswick(罗格斯大学新布朗斯维尔回声分校计算机科学系)
;
The Hong Kong University of Science and Technology (GZ)(香港科学与技术大学(GZ))
;
Shanghai AI Laboratory(上海人工智能实验室)
SignScene: Visual Sign Grounding for Mapless Navigation
SignScene: 用于无地图导航的视觉标志接地
Nicky Zimmerman, Joel Loo, Benjamin Koh, Zishuo Wang, David Hsu
机构
*
Smart Systems Institute, National University of Singapore, 3 Research Link, 117602, Singapore.(新加坡国立大学智能系统研究所)
;
School of Computing, National University of Singapore, 13 Computing Drive, 117417, Singapore.(新加坡国立大学计算机学院)
Advancing Multi-Robot Networks via MLLM-Driven Sensing, Communication, and Computation: A Comprehensive Survey
通过MLLM驱动的感知、通信与计算推进多机器人网络:综述
Hyun Jong Yang, Howon Lee, Kyuhong Shim, Jeongho Kwak, Hyunsoo Kim, Donghoon Kim, Khoa Anh Ngo, Sehyun Ryu, Jaehyun Choi, Youbin Kim, Chanjun Moon, Michael Ryoo, Byonghyo Shim
机构
*
Seoul National University(首尔大学)
;
Ajou University(亚洲大学)
;
Sungkyunkwan University(成均馆大学)
;
Korea University(高丽大学)
;
POSTECH(浦项科技大学)
;
Stony Brook University(石溪大学)
专题命中
视觉定位与Grounding
:MLLM(title,abstract);grounding(abstract);multimodal large language model(abstract)
机构
*
National University of Singapore(新加坡国立大学)
;
Fudan University(复旦大学)
;
Tsinghua University(清华大学)
;
Zhejiang University(浙江大学)
;
University of Science and Technology of China(中国科学技术大学)
;
vivo
SafeHumanoid: VLM-RAG-driven Control of Upper Body Impedance for Humanoid Robot
SafeHumanoid: 通过VLM-RAG驱动的人形机器人上半身阻抗控制
Yara Mahmoud, Jeffrin Sam, Nguyen Khang, Marcelino Fernando, Issatay Tokmurziyev, Miguel Altamirano Cabrera, Muhammad Haris Khan, Artem Lykov, Dzmitry Tsetserukou
机构
*
Skolkovo Institute of Science and Technology(斯克洛尔沃科学与技术研究所)
专题命中
视觉定位与Grounding
:VLM(title,abstract);vision language model(abstract);grounding(abstract)
机构
*
Department of Electrical and Computer Engineering, The University of Alabama(电气与计算机工程系,阿拉巴马大学)
;
Electrical and Computer Engineering, Michigan State University(电气与计算机工程,密歇根州立大学)
;
Agricultural and Biological Engineering, Mississippi State University(农业与生物工程,密苏里州立大学)
;
Department of Computer Science, University of Alabama(计算机科学系,阿拉巴马大学)
;
USDA-ARS Genetics and Sustainbale Agriculture(美国农业部ARS基因与可持续农业)