CommentsThis document is the unedited Author's version of a Submitted Work to Computers and Chemical Engineering. 19 pages, 2 figures, SI available (15 pages)
Cross-modal Context-aware Learning for Visual Prompt Guided Multimodal Image Understanding in Remote Sensing
跨模态上下文感知学习:用于遥感中视觉提示引导的多模态图像理解
Xu Zhang, Jiabin Fang, Zhuoming Ding, Jin Yuan, Xuan Liu, Qianjun Zhang, Zhiyong Li
机构
*
College of Computer Science and Electronic Engineering, Hunan University(计算机科学与电子工程学院,湖南大学)
;
School of Robotics and the National Engineering Research Center of Robot Visual Perception and Control Technology, Hunan University(机器人学院及机器人视觉感知与控制技术国家工程研究中心,湖南大学)
;
School of Computing and Artificial Intelligence, Southwest Jiaotong University(计算与人工智能学院,西南交通大学)
Guard Me If You Know Me: Protecting Specific Face-Identity from Deepfakes
Kaiqing Lin, Zhiyuan Yan, Ke-Yue Zhang, Li Hao, Yue Zhou, Yuzhen Lin, Weixiang Li, Taiping Yao, Shouhong Ding, Bin Li
机构
*
Guangdong Provincial Key Laboratory of Intelligent Information Processing(广东省智能信息处理重点实验室)
;
Shenzhen Key Laboratory of Media Security(深圳媒体安全重点实验室)
;
SZU-AFS Joint Innovation Center for AI Technology(深圳大学-腾讯优图联合创新中心)
;
Shenzhen University(深圳大学)
;
School of Electronic and Computer Engineering(电子与计算机工程学院)
;
Peking Univerisity(北京大学)
;
Tencent Youtu Lab(腾讯优图实验室)