机构
*
University of Science and Technology of China(中国科学技术大学)
;
Shanghai Advanced Research Institute, Chinese Academy of Sciences(上海先进研究院,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
OmniVTG: A Large-Scale Dataset and Training Paradigm for Open-World Video Temporal Grounding
OmniVTG:一种大规模数据集和开放世界视频时间定位的训练范式
Minghang Zheng, Zihao Yin, Yi Yang, Yuxin Peng, Yang Liu
机构
*
Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所)
;
State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室)
;
Central Media Technology Institute, Huawei Technologies Ltd.(华为技术有限公司中央媒体技术研究所)
;
PKU-WUHAN Institute for Artificial Intelligence, Peking University(北京大学武汉人工智能研究所)
Hard to See, Hard to Label: Generative and Symbolic Acquisition for Subtle Visual Phenomena
难以察觉,难以标注:生成与符号获取用于微妙的视觉现象
Renjith Prasad, Rishabh Sharma, Andrew E. Shao, Annmary Justine Koomthanam, Shreyas Kulkarni, Suparna Bhattacharya, Martin Foltin, Amit Sheth, David Orozco, Matthew Quinn, Brian Sammuli
机构
*
University of South Carolina(南卡罗来纳大学)
;
AI Research Lab, HPE Labs, Hewlett Packard Enterprise(HPE实验室人工智能研究部)
;
Indian AI Research Organization(印度人工智能研究组织)
;
General Atomics(通用原子公司)
Gesture2Music: A Low-Latency Real-Time Framework for Continuous Gesture-Driven Music Generation
Gesture2Music: 一种低延迟的实时框架,用于连续的手势驱动音乐生成
Rathinaraja Jeyaraj, Barathi Subramanian, Kapilya Gangadharan, Anand Paul
机构
*
Stanford University(斯坦福大学)
;
Saveetha Institute of Medical and Technical Sciences(Saveetha医学与技术科学学院)
;
LSU Health Sciences Center New Orleans(路易斯安那州立大学健康科学中心新奥尔良分校)
High-Precision Dichotomous Image Segmentation via Depth Integrity-Prior and Fine-Grained Patch Strategy
通过深度完整性先验和细粒度补丁策略实现高精度二元图像分割
Xianjie Liu, Keren Fu, Qijun Zhao
机构
*
College of Computer Science, Sichuan University(四川大学计算机学院)
;
National Key Lab of Fundamental Science on Synthetic Vision, Sichuan University(合成视觉基础科学国家实验室)