Personal Care Utility (PCU): Building the Health Infrastructure for Everyday Insight and Guidance
机构 * University of California, Irvine(加州大学尔湾分校)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CL、cs.AI
Comments 22 pages, 2 figures, 1 table, Journal paper
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * University of California, Irvine(加州大学尔湾分校)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CL、cs.AI
Comments 22 pages, 2 figures, 1 table, Journal paper
机构 * LGND AI ; Tolbi
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV、cs.AI
Comments 27 pages, 7 figures
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; Zhejiang University(浙江大学) ; The University of Tokyo(东京大学) ; Fudan University(复旦大学) ; Nanjing University(南京大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted at NeurIPS 2025
机构 * Southeast University(东南大学) ; Monash University(墨尔本大学) ; Xiaohongshu Inc.(小红书公司) ; University of Southern California(南加州大学) ; Fudan University(复旦大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * University of Science and Technology of China(中国科学技术大学) ; Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; Beihang University(北京航空航天大学) ; Shanghai Jiao Tong University(上海交通大学) ; Zhejiang University(浙江大学) ; State Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室) ; Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究所)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * University of Bristol(布里斯托大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
机构 * Sber AI Lab(Sber AI实验室) ; Samara State Medical University(萨马拉州医学大学) ; ISP RAS Research Center for Trusted Artificial Intelligence(俄罗斯科学院信息与通信技术研究所可信人工智能研究中心)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted to ACMMM 2025, Datasets track
机构 * Hamlyn Centre for Robotic Surgery, Institute of Global Health Innovation, Imperial College London(帝国理工学院伦敦校区全球健康创新研究所机器人手术中心) ; Department of Mechanical Engineering and Automation, Fuzhou University(福州大学机械工程与自动化学院) ; TAMS Group, Informatics, University of Hamburg(汉堡大学信息学院TAMS集团)
专题命中 视频多模态 :multi-modal(abstract)
Comments 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
机构 * City University of Hong Kong(香港城市大学) ; Tencent(腾讯) ; Zhejiang University(浙江大学)
专题命中 跨模态检索 :multimodal(title,abstract);MLLM(title,abstract);分类 cs.CV
Comments NeurIPS 2025
机构 * The Pennsylvania State University(宾夕法尼亚州立大学) ; Intel(英特尔)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV、cs.CL
Comments Accepted at NeurIPS 2025 UniReps Workshop
机构 * Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(香港科学与技术大学计算机科学与工程系) ; College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) ; Department of Radiology, Union Hospital, Tongji Medical College, Huazhong University of Science and Technology(华中科技大学同济医学院附属同济医院放射科) ; Department of Pathology, Sir Run Run Shaw Hospital, School of Medicine, Zhejiang University(浙江大学医学院附属邵氏医院病理科) ; Department of Pathology, The Central Hospital of Wuhan, Tongji Medical College, Huazhong University of Science and Technology(华中科技大学同济医学院附属武汉中心医院病理科) ; Department of Pathology, The First Affiliated Hospital, School of Medicine, Zhejiang University(浙江大学医学院附属第一医院病理科) ; Department of Chemical and Biological Engineering, The Hong Kong University of Science and Technology(香港科学与技术大学化学与生物工程系) ; Division of Life Science, The Hong Kong University of Science and Technology(香港科学与技术大学生命科学系)
专题命中 跨模态检索 :multimodal(title);multi-modal(abstract);分类 cs.CV
机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) ; Washington State University(华盛顿州立大学) ; VinUniversity(文大学) ; Griffin University(格里芬大学) ; Hanoi University of Science and Technology(河内科学技术大学)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.AI
Comments Accepted at NeurIPS 2025
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.MM
Comments We identified that our current approach achieves its reported performance only under specific data conditions, and its robustness is weaker than we initially expected
机构 * Computer Science, Louisiana State University(计算机科学,路易斯安那州立大学) ; Computer Science, Furman University(计算机科学,福兰明大学) ; Environmental Sciences and Center for Computation and Technology, Louisiana State University(环境科学与计算技术中心,路易斯安那州立大学)
专题命中 跨模态检索 :multimodal(title,abstract)
机构 * Facultad de Ingeniería y Ciencias Básicas, Universidad Autónoma de Occidente(工程与基础科学学院,自治大学)
专题命中 跨模态检索 :multimodal(abstract);分类 cs.AI
Comments pages, 1 figure, 3 tables. Preprint. Evaluated on UCI Heart Disease (1989) and UCI Differentiated Thyroid Cancer Recurrence (2023). Uses IEEEtran
机构 * Department of Computer Science(计算机科学系) ; Columbia University(哥伦比亚大学) ; College of Medicine(医学院) ; SUNY Downstate Health Sciences University(SUNY 下州健康科学大学) ; Department of Mechanical Engineering(机械工程系)
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV
机构 * The University of Sydney(悉尼大学)
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.AI
Comments Prepared as part of the CERN openlab programme 2025. Also available on Zenodo, a repository operated by CERN and co-funded by the European Union
专题命中 跨模态检索 :multi-modal(abstract)
专题命中 跨模态检索 :multimodal(abstract)
Comments 32 pages, 6 figures
机构 * School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院) ; Zhongguancun Academy(中关村学院) ; Shanghai Innovation Institute(上海创新研究院) ; Lehigh University(莱特大学)
专题命中 多模态生成 :multi-modal(title,abstract);MLLM(abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted by NeurIPS 2025
机构 * School of Information Science and Engineering, Yunnan University(云南大学信息科学与工程学院) ; National University of Singapore(新加坡国立大学) ; Engineering Research Center of Cyberspace, Yunnan University(云南大学网络空间研究院)
专题命中 多模态生成 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV、cs.MM
机构 * DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团)
专题命中 多模态生成 :multi-modal(title);multimodal(abstract);分类 cs.CV、cs.CL
Comments 15 pages, 5 figures
机构 * Shandong University(山东大学) ; National University of Singapore(新加坡国立大学) ; Kuaishou Technology(快手科技) ; Harbin Institute of Technology(哈尔滨工业大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments Preprint
机构 * Department of Geography(地理系) ; University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
专题命中 多模态生成 :multi-modal(title);multimodal(abstract);分类 cs.CV
Comments 41 pages
专题命中 多模态生成 :multimodal(title,comments);分类 cs.CV、cs.CL、cs.AI
Comments 24 pages, 10 figures, 11 tables. Code can be found at: https://github.com/SinghNayanKumar/multimodal-product-lister/
机构 * Adobe Research(Adobe研究机构) ; University of Michigan(密歇根大学) ; UNC Chapel Hill(北卡罗来纳大学教堂山分校)
专题命中 多模态生成 :multimodal(abstract);MLLM(abstract);分类 cs.CV、cs.CL、cs.AI
Comments ICCV 2025; First three authors contributed equally. Project page: https://veggie-gen.github.io/
机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) ; The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; University of Chicago(芝加哥大学) ; Stanford University(斯坦福大学)
专题命中 多模态生成 :cross-modal(title);分类 cs.CV
机构 * Wangxuan Institute of Computer Technology, Peking University(计算机技术研究所,北京大学)
专题命中 多模态生成 :image-text(abstract);分类 cs.CV、cs.CL
机构 * Technical University of Munich(慕尼黑技术大学) ; ETH Zurich(苏黎世联邦理工学院)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
专题命中 多模态生成 :image-text(abstract);分类 cs.AI
Comments Finding unfinished issue in this work , still refining