Dingtalk DeepResearch: A Unified Multi Agent Framework for Adaptive Intelligence in Enterprise Environments
机构 * Industrial Brain Team, Dingtalk, Alibaba Group(钉钉工业大脑团队,钉钉,阿里巴巴集团)
专题命中 多模态Agent :multimodal(abstract);分类 cs.CL、cs.AI
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Industrial Brain Team, Dingtalk, Alibaba Group(钉钉工业大脑团队,钉钉,阿里巴巴集团)
专题命中 多模态Agent :multimodal(abstract);分类 cs.CL、cs.AI
专题命中 多模态Agent :multi-modal(abstract)
机构 * Chongqing University of Posts and Telecommunications(重庆邮电大学) ; Xi’an Jiaotong University(西安交通大学) ; University of New South Wales(新南威尔士大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(title,abstract);分类 cs.CV、cs.MM
Comments The paper will be published in the MMAsia2025 conference proceedings
机构 * Computer Science Department, Carnegie Mellon University(卡内基梅隆大学计算机科学系) ; Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所) ; Neuroscience Institute, Carnegie Mellon University(卡内基梅隆大学神经科学研究所) ; Department of Psychology, The University of Texas at Austin(德克萨斯大学奥斯汀分校心理学系) ; Center for Perceptual Systems, The University of Texas at Austin(德克萨斯大学奥斯汀分校感知系统中心)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL、cs.AI
Comments 22 pages, first three authors equal contribution
机构 * School of Control Science and Engineering, Shandong University(控制科学与工程学院,山东大学) ; Engineering Research Center of Intelligent Unmanned System, Ministry of Education(智能无人机系统工程研究中心,教育部) ; Department of Psychiatry, The Chinese University of Hong Kong(心理学系,香港中文大学) ; Department of Electronic Engineering, The Chinese University of Hong Kong(电子工程系,香港中文大学) ; AICARE Lab, Guangdong Medical University(AICARE实验室,广东医科大学) ; Department of Neurology, Shandong University of Traditional Chinese Medicine Affiliated Hospital(神经内科,山东中医药大学附属医院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments 35 pages, 8 figures, and 7 tables
机构 * Holcombe Department of ECE(霍尔科姆电气与计算机工程系) ; Clemson University(克莱姆森大学) ; University of Arizona(亚利桑那大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV
机构 * College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院) ; Shanghai Innovation Institute(上海创新研究院) ; Shanghai AI Lab(上海人工智能实验室) ; Hunyuan, Tencent(腾讯 Hunyuan)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV
Comments [NeurIPS2025] Project Page: https://codegoat24.github.io/UnifiedReward/think
机构 * Department of Computer Science & Engineering, Chalmers University of Technology and University of Gothenburg(计算机科学与工程系,楚姆勒斯技术大学和哥德堡大学)
专题命中 多模态训练与对齐 :image-text(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments After the publication of the paper, we discovered some significant errors/omissions that need to be corrected and improved
机构 * Fujian Key Laboratory of Sensing and Computing for Smart Cities, Xiamen University, China(福建智能城市感知与计算重点实验室,厦门大学,中国) ; Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University, China(多媒体可信感知与高效计算重点实验室,中华人民共和国教育部,厦门大学,中国) ; University of Science and Technology of China, China(中国科学技术大学,中国)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
Comments 17 pages, 7 figures, NeurIPS 2025
机构 * Wuhan University(武汉大学) ; Kuaishou Technology(快手科技)
专题命中 其他多模态 :multi-modal(abstract);分类 cs.AI
Comments to be published in NeurIPS 2025
专题命中 其他多模态 :multimodal(abstract)
Comments 5 pages, 4 figures, Supplementary material. In v2: streamlined explanations in the main text and added appendices about global spectrum and counting statistics; In v3: added DOI