Ethical Medical Image Synthesis
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
机构 * Stony Brook University(石溪大学) ; University of Texas at Austin(德克萨斯大学奥斯汀分校) ; University of Maryland(马里兰大学)
专题命中 视觉定位与Grounding :visual language model(abstract);分类 cs.CV
机构 * West Virginia University(西弗吉尼亚大学) ; University of Aberdeen(阿伯丁大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments 3rd Workshop in Data Engineering in Medical Imaging (DEMI), MICCAI-2025 Workshop
机构 * University of Science, VNU-HCM(越南胡志明市科学大学) ; University of Information Technology, VNU-HCM(越南胡志明市信息技术大学) ; Vietnam National University(越南国家大学) ; University of Dayton(戴维森大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
机构 * College of Computer Science and Technology, Zhejiang University, China(浙江大学计算机科学与技术学院) ; College of Software Technology, Zhejiang University, China(浙江大学软件学院) ; School of Geological Engineering and Geomatics, Chang’an University, China(长安大学地质工程与测绘学院) ; College of Biosystems Engineering and Food Science, Zhejiang University, China(浙江大学生物系统工程与食品科学学院) ; Faculty of Computer and Information Sciences, Hosei University, Japan(立命馆大学计算机与信息科学系)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
机构 * Monash University(墨尔本大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted to ACM ICMI 2025 Demos
机构 * Fan Yang
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
机构 * School of Physics and Optoelectronic Engineering(物理与光电工程学院) ; Guangdong-HongKong-Macao Joint Laboratory for Intelligent Micro-Nano Optoelectronic Technology(粤港澳联合智能微纳光电技术实验室) ; School of Information Engineering and Automation(信息工程与自动化学院)
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.CV
机构 * Department of Computer Science University of Basel Basel, Switzerland(计算机科学系 巴塞尔大学 巴塞尔瑞士)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG
机构 * Department of Computer Science(计算机科学系) ; Virginia Tech(弗吉尼亚理工大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
机构 * RMIT University, Australia(皇家墨尔本理工大学)
专题命中 视觉定位与Grounding :vision language model(abstract);分类 cs.CV
Comments 19 Pages, 24 Figures
机构 * LITIV, Polytechnique Montréal(Polytechnique Montréal 的 LITIV) ; Data Science Laboratory, Université du Québec (TELUQ)(Université du Québec (TELUQ) 的 Data Science Laboratory)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments 10 pages
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Project page: https://github.com/jasongief/TGS-Agent
机构 * College of Intelligence and Computing, Tianjin University(智能与计算学院,天津大学) ; Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院) ; Shenzhen University of Advanced Technology(深圳大学先进技术学院)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
机构 * Rensselaer Polytechnic Institute(伦斯勒理工学院) ; Princeton University(普林斯顿大学) ; Cisco Research(思科研究) ; Cornell University(康奈尔大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG
Comments Cameral-ready version: added experiments using the HPSv2 reward, improved notation consistency for the diffusion model, and added related works
机构 * Department of Computer Science, George Mason University(乔治·马歇尔大学计算机科学系) ; Department of Computer Science, The University of Texas at Austin(德克萨斯大学奥斯汀分校计算机科学系) ; Sony AI(索尼人工智能)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments 7 pages, IROS 2025
机构 * University of California, San Diego(加州大学圣地亚哥分校)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
机构 * Centre for Intelligent Sensing, Queen Mary University of London(智能传感中心,伦敦女王玛丽大学) ; Idiap Research Institute and École Polytechnique Fédérale de Lausanne(Idiap研究机构和日内瓦联邦理工学院)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments 7 pages,18 figures,2 tables
机构 * School of Automation, Beijing Institute of Technology, Beijing, China(自动化学院,北京理工大学)
专题命中 视觉定位与Grounding :MLLM(abstract);分类 cs.CV
Comments IROS2025
机构 * Sun Yat-sen University(中山大学) ; ByteDance Intelligent Creation(字节跳动智能创作) ; University of Surrey(Surrey大学) ; Alibaba Cloud Computing(阿里巴巴云计算)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted to ICCV 2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
机构 * University of Washington(华盛顿大学) ; Amazon Lab126(亚马逊实验室126) ; National Tsing Hua University(国立清华大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments ICCV 2025
机构 * Fudan University, China(复旦大学)
专题命中 视觉定位与Grounding :MLLM(abstract);分类 cs.CV
Comments ICCV 2025, Project Page: https://henghuiding.com/OmniAVS/
机构 * Meituan Inc.(美团公司) ; Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Accepted by ICCV 2025
机构 * University Of Houston(休斯顿大学) ; Mysten Labs(Mysten实验室) ; Kent State University(肯特州立大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
机构 * CISPA Helmholtz Center for Information Security(CISPA 欧洲信息安全研究中心) ; TU Delft(代尔夫特理工大学)
专题命中 视觉定位与Grounding :vision language model(abstract);分类 cs.CV
Comments Accepted at ICCV 2025
机构 * eBay Inc(eBay公司)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG
Comments Accepted at SIGIR eCom'25. https://sigir-ecom.github.io/eCom25Papers/paper_23.pdf
机构 * Technical University of Munich(慕尼黑技术大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
机构 * Tianjin University(天津大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
机构 * Peking University(北京大学) ; Guangdong University of Technology(广东工业大学) ; The University of Sheffield(谢菲尔德大学) ; University of Science and Technology Beijing(北京科技大学) ; Tsinghua University(清华大学) ; The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; Nanjing University(南京大学) ; University of Trento(特伦特大学) ; Queen's University Belfast(贝尔法斯特女王大学)
专题命中 视觉定位与Grounding :MLLM(abstract);分类 cs.CV
Comments Paper was accepted by ACM MM 2025; Code: https://github.com/YihuaJerry/EventVAD