OralAgent: Integrating Reasoning, Tools, and Knowledge for Interactive Dental Image Analysis
OralAgent: 融合推理、工具与知识的交互式牙科影像分析
Jing Hao, Siyuan Dai, Yongxin Zhang, Yuci Liang, Jiamin Wu, Jiahao Bao, Yuxuan Fan, Zanting Ye, Yanpeng Sun, Xinyu Zhang, Ming Hu, Liang Zhan, James Kit Hon Tsoi, Linlin Shen, Junjun He, Kuo Feng Hung
机构
*
Faculty of Dentistry, the University of Hongkong, Hong Kong SAR, China(香港大学牙科学院,中国香港特别行政区)
;
Department of Electrical and Computer Engineering, University of Pittsburgh, Pittsburgh, PA, USA(匹兹堡大学电气与计算机工程系,美国宾夕法尼亚州匹兹堡)
;
Shenzhen University, China(深圳大学,中国)
;
Department of Craniomaxillofacial Surgery, Shanghai Ninth People’s Hospital, China(上海第九人民医院口腔颌面外科部,中国)
;
Nanyang technological University, Singapore(南洋理工大学,新加坡)
;
School of Biomedical Engineering, Southern Medical University, China(南方医科大学生物医学工程学院,中国)
;
Singapore University of Technology and Design, Singapore(新加坡科技设计大学,新加坡)
;
University of Auckland, new zealand(奥克兰大学,新西兰)
;
Shanghai Artificial Intelligence Laboratory , China(上海人工智能实验室,中国)
CommentsAdded a conceptual diagram for the LGC architecture, 14 pages, 10 figures, 7 tables. Submitted to IEEE Transactions on Information Forensics and Security. The source code is available at https://github.com/eihmuekhine/Latent-Geometric-Chords
机构
*
Key Laboratory of Target Cognition and Application Technology (TCAT), AIRI, CAS(目标认知与应用技术重点实验室(TCAT),空气动力研究所,中国科学院)
;
School of Electronic, Electrical and Communication Engineering, UCAS(电子、电气与通信工程学院,中国科学院大学)
;
Aerospace Information Research Institute, Chinese Academy of Sciences(航天信息研究所,中国科学院)
;
Baidu Inc.(百度公司)
VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation
VisRAG2.0:通过视觉检索增强生成中的证据引导多图像推理减轻视觉幻觉
Yubo Sun, Chunyi Peng, Yukun Yan, Shi Yu, Zhenghao Liu, Sen Mei, Chi Chen, Maosong Sun
机构
*
School of Software and Microelectronics, Peking University, China(北京大学软件与微电子学院)
;
School of Computer Science and Engineering, Northeastern University, China(东北大学计算机科学与工程学院)
;
Department of Computer Science and Technology, Institute for AI, Tsinghua University, China(清华大学人工智能研究院计算机科学与技术系)
机构
*
Chongqing University(重庆大学)
;
Independent Researcher(独立研究者)
;
University of the Chinese Academy of Sciences(中国科学院大学)
;
Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航天信息研究所)
Alexandre Sallinen, Stefan Krsteski, Paul Teiletche, Marc-Antoine Allard, Baptiste Lecoeur, Michael Zhang, Fabrice Nemo, David Kalajdzic, Matthias Meyer, Mary-Anne Hartley
机构
*
École Polytechnique Fédérale de Lausanne (EPFL), Switzerland(瑞士联邦理工学院洛桑校区)
;
ETH Zürich, Switzerland(瑞士苏黎世联邦理工学院)
;
T.H. Chan School of Public Health, Harvard University, USA(哈佛大学T.H. Chan公共卫生学院)
专题命中
多模态RAG
:RAG(title,abstract);分类 cs.AI
CommentsThis paper was originally submitted to the CODEML workshop for ICML 2025. 9 pages (including references and appendices)
CLAIR-Fin: An Adversarial Multi-Agent Framework for Claim-Level Verification and Adaptive Debate in Cross-Modal Financial QA
CLAIR-Fin:用于跨模态金融问答中声明级验证与自适应辩论的对抗性多智能体框架
Fatema Tuj Johora Faria, Mukaffi Bin Moin, Jubayer Al Mahmud, M. F. Mridha, Md. Alam Hossain
机构
*
Ahsanullah University of Science and Technology(阿萨努拉科技大学)
;
Jashore University of Science and Technology(杰索尔科技大学)
;
American International University - Bangladesh(孟加拉国美国国际大学)
Adaptive Multimodal Agents-Based Framework for Automatic Workflow Execution
基于自适应多智能体框架的自动工作流执行
Susanna Cifani, Mario Luca Bernardi, Marta Cimitile
机构
*
Sapienza University of Rome(罗马萨皮恩扎大学)
;
Department of Engineering University of Sannio(萨尼奥大学工程系)
;
Faculty of Jurisprudence Unitelma Sapienza University(法理学院萨皮恩扎大学)
CommentsCopyright 2026 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses. Accepted for publication at the 2026 IEEE International Conference on Evolving and Adaptive Intelligent Systems (EAIS 2026)
M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language Models
M3DocDep: 多模态、多页、多文档依赖分块方法基于大视觉-语言模型
Joongmin Shin, Jeongbae Park, Jaehyung Seo, Heuiseok Lim
机构
*
Human-inspired AI Research, Korea University(韩国大学人智AI研究所)
;
Computer Science and Engineering, Konkuk University(konkuk大学计算机科学与工程系)
;
Department of Computer Science and Engineering, Korea University(韩国大学计算机科学与工程系)
AI Slop or AI-enhancement? Student perceptions of AI-generated media for an English for Academic Purposes course
AI 产出物还是AI增强?英语学术用途课程中学生对AI生成媒体的看法
David James Woo, Deliang Wang, Kai Guo
机构
*
Everwrite Limited(Everwrite有限公司)
;
Faculty of Education, The University of Hong Kong(香港大学教育学院)
;
Faculty of Education, The Chinese University of Hong Kong(香港中文大学教育学院)