ACM Multimedia Grand Challenge on ENT Endoscopy Analysis
Trong-Thuan Nguyen, Viet-Tham Huynh, Thao Thi Phuong Dao, Ha Nguyen Thi, Tien To Vu Thuy, Uyen Hanh Tran, Tam V. Nguyen, Thanh Dinh Le, Minh-Triet Tran
机构
*
University of Science, VNU-HCM(越南胡志明市科学大学)
;
Vietnam National University(越南国家大学)
;
Thong Nhat Hospital(通那医院)
;
Faculty of Medicine, Pham Ngoc Thach University of Medicine(范 Ngoc Thach 医学院)
;
Cho Ray Hospital(Cho Ray 医院)
;
University of Dayton(Dayton 大学)
;
University of Health Sciences, VNU-HCM(越南胡志明市健康科学大学)
Whose Truth? Pluralistic Geo-Alignment for (Agentic) AI
Krzysztof Janowicz, Zilong Liu, Gengchen Mai, Zhangyu Wang, Ivan Majic, Alexandra Fortacz, Grant McKenzie, Song Gao
机构
*
University of Vienna(维也纳大学)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
University of Maine(缅因大学)
;
McGill University(麦吉尔大学)
;
University of Wisconsin(威斯康星大学)
Magic Fixup: Streamlining Photo Editing by Watching Dynamic Videos
Hadi Alzayer, Zhihao Xia, Xuaner Zhang, Eli Shechtman, Jia-Bin Huang, Michael Gharbi
机构
*
University of Maryland \& Adobe College Park MD USA
;
Adobe San Jose CA USA
;
Adobe Seattle CA USA
;
Adobe San Francisco CA USA
;
University of Maryland \& Adobe
;
Adobe
EarthSynth: Generating Informative Earth Observation with Diffusion Models
Jiancheng Pan, Shiye Lei, Yuqian Fu, Jiahao Li, Yanxing Liu, Yuze Sun, Xiao He, Long Peng, Xiaomeng Huang, Bo Zhao
机构
*
Tsinghua University(清华大学)
;
University of Sydney(悉尼大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Wuhan University(武汉大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Shanghai Jiao Tong University(上海交通大学)
MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space
Lixing Xiao, Shunlin Lu, Huaijin Pi, Ke Fan, Liang Pan, Yueer Zhou, Ziyong Feng, Xiaowei Zhou, Sida Peng, Jingbo Wang
机构
*
Zhejiang University(浙江大学)
;
The Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳))
;
The University of Hong Kong(香港大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
DeepGlint
;
Shanghai AI Laboratory(上海人工智能实验室)
RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers
Min Zhao, Guande He, Yixiao Chen, Hongzhou Zhu, Chongxuan Li, Jun Zhu
机构
*
Dept. of Comp. Sci. \& Tech., BNRist Center, THU-Bosch ML Center, Tsinghua University.
;
The University of Texas at Austin.
;
Gaoling School of Artificial Intelligence Renmin University of China Beijing, China.
;
Beijing Key Laboratory of Research on Large Models
;
Engineering Research Center of Next-Generation Intelligent Search
;
Pazhou Laboratory (Huangpu)
机构
*
Division of Robotics, Perception and Learning, KTH Royal Institude of Technology(机器人、感知与学习系,皇家理工学院)
;
Department of Computer Science, University of Copenhagen(计算机科学系,哥本哈根大学)
专题命中
扩散模型
:diffusion(title,abstract)
Comments\c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works
机构
*
Peking University(北京大学)
;
Wuhan University(武汉大学)
;
Shanghai University of International Business(上海国际商务大学)
;
Microsoft(微软公司)
专题命中
扩散模型
:diffusion(title,abstract)
CommentsSIGIR 2025
Journal refSIGIR 2025: Proceedings of the 48th International ACM SIGIR Conference on Research and Development in Information Retrieval Pages 1593 - 1602
Estimating Musical Surprisal from Audio in Autoregressive Diffusion Model Noise Spaces
Mathias Rose Bjare, Stefan Lattner, Gerhard Widmer
专题命中
扩散模型
:diffusion(title,abstract)
Comments9 pages, 1 figure, 5 tables. Accepted at the 25th International Society for Music Information Retrieval Conference (ISMIR), Daejeon, South Korea, 2025 2025
Well-Posedness of the Cauchy Problem for One-Dimensional Nonlinear Diffusion Equations with Dynamic and Fourth-Type Boundary Conditions in the Lp Lq Maximal Regularity Setting
MagicHOI: Leveraging 3D Priors for Accurate Hand-object Reconstruction from Short Monocular Video Clips
Shibo Wang, Haonan He, Maria Parelli, Christoph Gebhardt, Zicong Fan, Jie Song
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
The Hong Kong University of Science and Technology(香港科技大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
Max Planck Institute for Intelligent Systems(智能系统马克斯·普朗克研究所)