机构
*
The Chinese University of Hong Kong Shanghai AI Laboratory(香港中文大学上海人工智能实验室)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai Innovation Institute(上海创新研究院)
A Versatile Diffusion Transformer with Mixture of Noise Levels for Audiovisual Generation
Gwanghyun Kim, Alonso Martinez, Yu-Chuan Su, Brendan Jou, José Lezama, Agrim Gupta, Lijun Yu, Lu Jiang, Aren Jansen, Jacob Walker, Krishna Somandepalli
机构
*
Seoul National University(首尔国立大学)
;
Google DeepMind(谷歌DeepMind)
;
Stanford University(斯坦福大学)
;
Carnegie Mellon University(卡内基梅隆大学)
ControlRadio: Prompt-Driven Controllable Diffusion for Cross-Modal Radio Map Generation
ControlRadio:用于跨模态无线电地图生成的提示驱动可控扩散模型
Kangjun Liu, Xiying Pan, Shuhang Zhang, Xiang Xiang, Ke Chen, Yaowei Wang
机构
*
Pengcheng Laboratory(鹏城实验室)
;
South China University of Technology(华南理工大学)
;
Peking University(北京大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
KVAE: Family of Tokenizers for Multimodal Generative Models
KVAE:用于多模态生成模型的分词器家族
Andrey Shutkin, Denis Parkhomenko, Ivan Kirillov, Kirill Chernyshev, Kirill Malakhov, Ilia Vasiliev, Ilia Trushkin, Valeriya Kobenko, David Chikovani, Alexander Ivanov, Azat Saginbaev, Egor Silvestrov, Ivan Mikheev, Konstantin Zakharov
Driving, Fast or Slow? Neuro-Symbolic Guidance for Motion Prediction in Multi-Modal Ground Mobility
驾驶,快或慢?多模态地面移动中运动预测的神经符号引导
Simon Kohaut, Felix Divo, Julius Hahnewald, Benedict Flade, Julian Eggert, Kristian Kersting, Devendra Singh Dhami
机构
*
Artificial Intelligence and Machine Learning Lab, TU Darmstadt(达姆施塔特工业大学人工智能与机器学习实验室)
;
Honda Research Institute(本田研究所)
;
Hessian Center for AI (hessian.AI)(黑森州人工智能中心)
;
Centre for Cognitive Science(认知科学中心)
;
German Center for AI (DFKI)(德国人工智能研究中心)
;
Uncertainty in Artificial Intelligence Lab, TU Eindhoven(埃因霍温理工大学人工智能不确定性实验室)