European Conference on Computer Vision · 会议 · Computer Vision
共收录 4781 篇
2606.199442026-06-19cs.CV新提交
Timage: A Generative Text-in-Image Paradigm for Fine-Tuning Vision-Language Models
Timage: 一种用于微调视觉语言模型的文本嵌入图像生成范式
Yifeng Wu, Huimin Huang, Ruiluo Wu, Chunyi Lin, Guanhua Chen, Xian Wu, Wang Song, Ruize Han
机构
*
Fudan University(复旦大学)
;
Shenzhen University of Advanced Technology(深圳先进技术大学)
;
Tencent Jarvis Lab(腾讯贾维斯实验室)
;
Southern University of Science and Technology(南方科技大学)
Counterfactual Intervention Feature Transfer for Visible-Infrared Person Re-identification
反事实干预特征迁移用于可见光-红外行人重识别
Xulin Li, Yan Lu, Bin Liu, Yating Liu, Guojun Yin, Qi Chu, Jinyang Huang, Feng Zhu, Rui Zhao, Nenghai Yu
机构
*
School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学)
;
Key Laboratory of Electromagnetic Space Information, Chinese Academy of Science(电磁空间信息重点实验室,中国科学院)
;
School of Data Science, University of Science and Technology of China(数据科学学院,中国科学技术大学)
;
SenseTime Research(商汤科技研究院)
;
Qing Yuan Research Institute, Shanghai Jiao Tong University(青元研究院,上海交通大学)
Rethinking Weakly-supervised Video Temporal Grounding From a Game Perspective
从博弈视角重新思考弱监督视频时间定位
Xiang Fang, Zeyu Xiong, Wanlong Fang, Xiaoye Qu, Chen Chen, Jianfeng Dong, Keke Tang, Pan Zhou, Yu Cheng, Daizong Liu
机构
*
Hubei Key Laboratory of Distributed System Security(湖北分布式系统安全重点实验室)
;
Hubei Engineering Research Center on Big Data Security(大数据安全工程研究中心)
;
School of Cyber Science and Engineering(网络安全科学与工程学院)
;
Huazhong University of Science and Technology(华中科技大学)
;
University of Central Florida(佛罗里达中央大学)
;
Zhejiang Gongshang University(浙江工商大学)
;
Guangzhou University(广州大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Peking University(北京大学)
Comments16 pages including Appendix, 14 figures and 4 tables. Revised for clarity; updated terminology and abstract; added URLs to GitHub and Harvard Dataverse. Using ECCV template
CODER: Coupled Diversity-Sensitive Momentum Contrastive Learning for Image-Text Retrieval
CODER: 耦合多样性敏感动量对比学习用于图像-文本检索
Haoran Wang, Dongliang He, Wenhao Wu, Boyang Xia, Min Yang, Fu Li, Yunlong Yu, Zhong Ji, Errui Ding, Jingdong Wang
机构
*
Department of Computer Vision Technology (VIS), Baidu Inc., Beijing, China(百度公司计算机视觉技术部(VIS),北京,中国)
;
Key Lab of Intelligent Information Processing of Chinese Academy of Sciences (CAS), Institute of Computing Technology, CAS, Beijing, China(中国科学院智能信息处理重点实验室,中国科学院计算技术研究所,北京,中国)
;
College of Information Science & Electronic Engineering, Zhejiang University, Hangzhou, China(浙江大学信息与电子工程学院,杭州,中国)
;
School of Electrical & Information Engineering, Tianjin University, Tianjin, China(天津大学电气与信息工程学院,天津,中国)
;
The University of Sydney, Sydney, Australia(悉尼大学,悉尼,澳大利亚)
Solving the inverse problem of microscopy deconvolution with a residual Beylkin-Coifman-Rokhlin neural network
利用残差Beylkin-Coifman-Rokhlin神经网络求解显微成像反问题
Rui Li, Mikhail Kudryashev, Artur Yakimovich
机构
*
Center for Advanced Systems Understanding (CASUS)(先进系统理解中心)
;
Helmholtz-Zentrum Dresden-Rossendorf e. V. (HZDR)(德累斯顿-罗斯托克研究所)
;
Max Delbrück Center for Molecular Medicine in the Helmholtz Association(马克斯·德尔布吕克分子医学研究中心)
;
Institute of Medical Physics and Biophysics, Charite-Universitätsmedizin(医学物理与生物物理研究所)
;
Institute of Computer Science, University of Wrocław(沃林福大学计算机科学研究所)
GenVideoLens: Where LVLMs Fall Short in AI-Generated Video Detection?
GenVideoLens:在AI生成视频检测中LVLMs的局限性
Yueying Zou, Pei Pei Li, Zekun Li, Xinyu Guo, Xing Cui, Huaibo Huang, Ran He
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
University of California, Santa Barbara(加州大学圣巴巴拉分校)
;
Center for Research on Intelligent Perception and Computing, NLPR, Institute of Automation, Chinese Academy of Sciences(智能感知与计算中心、国家智能感知与信息处理实验室、中国科学院自动化研究所)
机构
*
Shandong Normal University(山东师范大学)
;
Qilu University of Technology(齐鲁工业大学)
;
Nanjing University of Science and Technology(南京理工大学)
;
National University of Defense Technology(国防科学技术大学)
;
Shandong University(山东大学)
CommentsAccepted by IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI). Journal version of arXiv:2409.00342 (ECCV 2024). Code is available at: https://github.com/LeapLabTHU/AdaGen
VRP-UDF: Towards Unbiased Learning of Unsigned Distance Functions from Multi-view Images with Volume Rendering Priors
VRP-UDF: 向多视图图像中通过体积渲染先验实现无偏的无符号距离函数学习
Wenyuan Zhang, Chunsheng Wang, Kanle Shi, Yu-Shen Liu, Zhizhong Han
机构
*
School of Software, Tsinghua University(清华大学软件学院)
;
China Telecom Wanwei Information Technology Co., Ltd.(中国电信万维信息技术有限公司)
;
Kuaishou Technology(快手技术)
;
Department of Computer Science, Wayne State University(韦恩州立大学计算机科学系)
Instance-dependent Noisy-label Learning with Graphical Model Based Noise-rate Estimation
实例依赖性噪声标签学习与基于图形模型的噪声率估计
Arpit Garg, Cuong Nguyen, Rafael Felix, Thanh-Toan Do, Gustavo Carneiro
机构
*
Australian Institute for Machine Learning, University of Adelaide, Australia(澳大利亚机器学习研究所,阿德莱德大学,澳大利亚)
;
Department of Data Science and AI, Monash University, Australia(数据科学与人工智能系,莫纳什大学,澳大利亚)
;
Centre for Vision, Speech and Signal Processing, University of Surrey, UK(视觉、语音与信号处理中心,萨里大学,英国)