Unbiased Object Detection Beyond Frequency with Visually Prompted Image Synthesis
超越频率的无偏目标检测:基于视觉提示的图像合成
Xinhao Cai, Liulei Li, Gensheng Pei, Tao Chen, Jinshan Pan, Yazhou Yao, Wenguan Wang
机构
*
Nanjing University of Science and Technology(南京理工大学)
;
Zhejiang University(浙江大学)
;
Department of Electrical and Computer Engineering, Sungkyunkwan University(成均馆大学电子与计算机工程系)
;
State Key Laboratory of Intelligent Manufacturing of Advanced Construction Machinery(先进施工机械智能制造国家重点实验室)
Token-Level Constraint Boundary Search for Jailbreaking Text-to-Image Models
针对文本到图像模型的令牌级约束边界搜索
Jiangtao Liu, Zhaoxin Wang, Handing Wang, Cong Tian, Yaochu Jin
机构
*
School of Artificial Intelligence, Xidian University(西安电子科技大学人工智能学院)
;
School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院)
;
Trustworthy and General Artificial Intelligence Laboratory, School of Engineering, Westlake University(之江实验室可信与通用人工智能实验室)
机构
*
Institute of Information Engineering, CAS(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
State Key Laboratory of AI Safety, Institute of Computing Technology, CAS(中国科学院计算技术研究所人工智能安全国家重点实验室)
;
School of Computer Science and Technology, University of Chinese Academy of Sciences(中国科学院大学计算机科学与技术学院)
;
BDKM, University of Chinese Academy of Sciences(中国科学院大学BDKM)
;
School of Computer Science and Technology, Beijing Institute of Technology(北京理工大学计算机科学与技术学院)
;
School of Cyber Science and Tech., Shenzhen Campus, Sun Yat-sen University(中山大学深圳校区网络安全科学与技术学院)
机构
*
Universiti Malaya(马来大学)
;
Monash University(莫纳什大学)
;
United Arab Emirates University(阿拉伯联合酋长国大学)
;
University of Science and Technology Beijing(北京科技大学)
;
Nanyang Technological University (NTU)(南洋理工大学)
;
City St George’s, University of London(伦敦大学圣乔治学院)
;
Augusta University(奥古斯塔大学)
;
Squirrel Ai Learning
DiCo: Disentangled Concept Representation for Text-to-image Person Re-identification
DiCo: 用于文本到图像人物重识别的解耦概念表示
Giyeol Kim, Chanho Eom
机构
*
organization= Department of Imaging Science, Graduate School of Advanced Imaging Science, Multimedia \& Film, Chung-Ang University , addressline= , city= Seoul , postcode= 06974 , state= , country= South Korea
;
organization= Department of Metaverse Convergence, Graduate School of Advanced Imaging Science, Multimedia \& Film, Chung-Ang University , addressline= , city= Seoul , postcode= 06974 , state= , country= South Korea
CulturalFrames: Assessing Cultural Expectation Alignment in Text-to-Image Models and Evaluation Metrics
CulturalFrames: 评估文本到图像模型与评估指标中的文化期望一致性
Shravan Nayak, Mehar Bhatia, Xiaofeng Zhang, Verena Rieser, Lisa Anne Hendricks, Sjoerd van Steenkiste, Yash Goyal, Karolina Stańczak, Aishwarya Agrawal
机构
*
Mila – Quebec AI Institute(魁北克AI研究院)
;
Université de Montréal(蒙特利尔大学)
;
McGill University(麦吉尔大学)
;
Google Research(谷歌研究)
;
Google DeepMind(谷歌DeepMind)
;
Samsung - SAIT AI Lab(三星-SAIT人工智能实验室)
;
ETH AI Center(苏黎世联邦理工学院人工智能中心)
机构
*
University of Washington, Seattle(华盛顿大学)
;
Bar-Ilan University(巴伊兰大学)
;
University of California, Irvine(加州大学伊文斯顿分校)
;
Allen Institute of AI(人工智能研究院)
Cross-modal Full-mode Fine-grained Alignment for Text-to-Image Person Retrieval
跨模态全模式细粒度对齐用于文本到图像人物检索
Hao Yin, Xin Man, Feiyu Chen, Jie Shao, Heng Tao Shen
机构
*
Shenzhen Institute for Advanced Study, University of Electronic Science and Technology of China(深圳先进研究所,电子科学与技术大学)
;
University of Electronic Science and Technology of China(电子科学与技术大学)
;
Sichuan Artificial Intelligence Research Institute(四川人工智能研究院)
DREAM: Scalable Red Teaming for Text-to-Image Generative Systems via Distribution Modeling
通过分布建模实现文本到图像生成系统可扩展的红队测试
Boheng Li, Junjie Wang, Yiming Li, Zhiyang Hu, Leyi Qi, Jianshuo Dong, Run Wang, Han Qiu, Zhan Qin, Tianwei Zhang
机构
*
School of Cyber Science(网络安全学院)
;
State Key Laboratory of Blockchain and Data Security(区块链与数据安全国家重点实验室)
;
Zhejiang University(浙江大学)
;
Tsinghua University(清华大学)