Toward Autonomous Laboratory Safety Monitoring with Vision Language Models: Learning to See Hazards Through Scene Structure
迈向自主实验室安全监控:通过场景结构学习识别危险
Trishna Chakraborty, Udita Ghosh, Aldair Ernesto Gongora, Ruben Glatt, Yue Dong, Jiachen Li, Amit K. Roy-Chowdhury, Chengyu Song
机构
*
Electrical and Computer Engineering, University of California, Riverside, USA(加州大学河滨分校电气与计算机工程系)
;
Computer Science and Engineering, University of California, Riverside, USA(加州大学河滨分校计算机科学与工程系)
;
Lawrence Livermore National Laboratory(劳伦斯利弗莫尔国家实验室)
机构
*
Key Lab of Optoelectronic Technology and Systems and Engineering Research Center of Industrial Computed Tomography Nondestructive Testing, Ministry of Education, Chongqing University(光电技术与系统国家重点实验室和工业计算机断层非破坏检测工程研究中心,教育部,重庆大学)
;
School of Biomedical Engineering, Sun Yat-sen University(中山大学生物医学工程学院)
;
Department of Radiology, Beijing Anzhen Hospital, Capital Medical University(北京安贞医院放射科,首都医科大学)
;
School of Information Engineering, Nanchang, Jiangxi(信息工程学院,南昌,江西)
;
Institute of High Performance Computing (IHPC), Agency for Science, Technology and Research (A*STAR)(高性能计算研究所(IHPC),科技研究局(A*STAR))
Physically Guided Visual Mass Estimation from a Single RGB Image
基于单个RGB图像的物理引导视觉质量估计
Sungjae Lee, Junhan Jeong, Yeonjoo Hong, Kwang In Kim
机构
*
Graduate School of Artificial Intelligence, POSTECH, South Korea(人工智能研究生院,POSTECH,韩国)
;
Department of Electrical Engineering, POSTECH, South Korea(电气工程系,POSTECH,韩国)
CommentsThis manuscript (arXiv:2503.03215) is being withdrawn at the supervisor's request. The content is preliminary and needs further internal revision and approval before public release. We will resubmit a revised version after completion. Apologies for the inconvenience
Venus: An Efficient Edge Memory-and-Retrieval System for VLM-based Online Video Understanding
Venus: 一种高效的边缘内存与检索系统用于基于VLM的在线视频理解
Shengyuan Ye, Bei Ouyang, Tianyi Qian, Liekang Zeng, Mu Yuan, Xiaowen Chu, Weijie Hong, Xu Chen
机构
*
School of Computer Science and Engineering, Sun Yat-sen University, Guangzhou, China(计算机科学与工程学院,中山大学,广州,中国)
;
Department of Information Engineering, The Chinese University of Hong Kong, Hong Kong SAR, China(信息工程系,香港中文大学,香港特别行政区,中国)
;
Shenzhen Smart City Communications Co., Ltd., China(深圳智慧城市通信有限公司,中国)
CommentsAccepted to NeurIPS 2025 Workshops: SPACE in Vision, Language, and Embodied AI; and What Makes a Good Video: Next Practices in Video Generation and Evaluation
CitySeeker: How Do VLMS Explore Embodied Urban Navigation With Implicit Human Needs?
CitySeeker: VLMS如何通过隐含人类需求探索具身城市导航
Siqi Wang, Chao Liang, Yunfan Gao, Erxin Yu, Sen Li, Yushi Li, Jing Li, Haofen Wang
机构
*
Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算机系)
;
Tongji University(同济大学)
;
Nanjing Institute of Geography and Limnology, Chinese Academy of Sciences(中国科学院南京地理与湖泊研究所)
;
Research Centre for Data Science & Artificial Intelligence(数据科学与人工智能研究中心)
AlignMerge - Alignment-Preserving Large Language Model Merging via Fisher-Guided Geometric Constraints
AlignMerge - 通过Fisher引导的几何约束实现的保持对齐的大语言模型合并
Aniruddha Roy, Jyoti Patel, Aman Chadha, Vinija Jain, Amitava Das
机构
*
AI Institute, University of South Carolina(AI研究院,南卡罗来纳大学)
;
Indian Institute of Technology, Kharagpur(印度理工学院,Khargpur分校)
;
Islamic University of Technology(伊斯兰科技大学)
;
Stanford University, USA(斯坦福大学,美国)
;
Amazon AI, USA(亚马逊AI,美国)
;
HCL(HCL公司)
;
Evalueserve(Evalueserve公司)
;
Apple (USA)(苹果(美国))
;
Google (USA)(谷歌(美国))
;
Pragya Lab, BITS Pilani, Goa(Pragya实验室, BITS Pilani,Goa分校)