A Multi-Biometrics for Twins Identification Based Speech and Ear
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.CV
专题命中 音频语音多模态 :multi-modal(abstract);分类 cs.CV
专题命中 音频语音多模态 :cross-modal(abstract);分类 eess.AS
Comments 5 pages, 2 figures
专题命中 音频语音多模态 :multi-modal(abstract);分类 cs.CL
专题命中 音频语音多模态 :audio-visual(abstract);分类 cs.CV
Comments Appears in: IEEE International Conference on Computer Vision (ICCV) 2017
专题命中 音频语音多模态 :multi-modal(abstract);分类 cs.CV
Comments Published In Proceedings of the 17th International Society for Music Information Retrieval Conference (2016)
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.AI
Comments The short version of this paper is accepted to appear as an abstract in the proceedings of AAAI-17 (student abstract and poster program)
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.CV
Comments 21 pages, 9 figures, 4 tables
Journal ref Computer Vision and Image Understanding, volume 153, December 2016, pages 64-76
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.CL
Comments Submission to Computer Speech and Language, special issue on Interaction Technologies for Children
专题命中 音频语音多模态 :multi-modal(abstract);分类 cs.CV
Comments The paper has been published on ICCCBE 2016. http://www.see.eng.osaka-u.ac.jp/seeit/icccbe2016/ http://www.see.eng.osaka-u.ac.jp/seeit/icccbe2016/download/Tentative_Time_Table_ICCCBE2016_2016-05-10.pdf
专题命中 音频语音多模态 :audio-visual(abstract);分类 cs.MM
Comments Proceedings of ACM Multimedia 2016
专题命中 音频语音多模态 :audio-visual(abstract);分类 cs.MM
Comments 15 pages, 8 figures
Journal ref IEEE Transactions on Audio, Speech, and Language Processing 23(4), 718-731, April, 2015
专题命中 音频语音多模态 :multi-modal(abstract);分类 cs.AI
Comments Please, find the supplementary video material at: http://sunai.uoc.edu/~vponcel/video/VOMSessionSample.mp4
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.CL
Comments 6 pages, accepted to NIPS 2015 Workshop on Machine Learning for Spoken Language Understanding and Interaction
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.AI
Comments ii + 20 pages
Journal ref SEKI Report SR-2008-01 (ISSN 1437-4447), Saarland University, 2008
专题命中 音频语音多模态 :audio-visual(abstract);分类 cs.CV
Comments Speech and Language Technologies (Book), Prof. Ivo Ipsic (Ed.), ISBN: 978-953-307-322-4, InTech (2011)
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.MM
Comments 2011 9th International Workshop on Content-Based Multimedia Indexing (CBMI), Madrid : Spain (2011)
专题命中 音频语音多模态 :audio-visual(abstract);分类 cs.CL
Journal ref In Proceedings of the EACL Workshop on Computational Linguistics for Literature, April 2014, Gothenburg, Sweden
专题命中 音频语音多模态 :multi-modal(abstract);分类 cs.AI
Journal ref Journal Of Artificial Intelligence Research, Volume 36, pages 71-128, 2009
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.CV
Comments 5 pages, 11 figures. arXiv admin note: text overlap with arXiv:1201.3720 and arXiv:1204.1177
Journal ref International Journal of Computer Applications 35(7):17-21, December 2011
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.CV
Comments Engineering for Human-Computer Interaction (EHCI'01),Toronto, Canada. May 11-14, 2001. Lecture Notes in Computer Science, Springer Verlag. 14 pages
专题命中 音频语音多模态 :multi-modal(abstract);分类 cs.CL
Comments 8 pages, 9 figures
Journal ref Proceedings of the Second International Conference on Language Resources and Evaluation, pp. 1699-1706, Paris: European Language Resources Association, 2000
机构 * HKUST(香港科技大学) ; MAP(多模态艺术投影)
专题命中 音频语音多模态 :分类 cs.AI、cs.MM、eess.AS;multimodal(comments)
机构 * Giant Network, China(中国巨网) ; Computer Science, University of Massachusetts Boston(马萨诸塞大学波士顿分校计算机科学系)
专题命中 音频语音多模态 :分类 cs.CV、cs.AI、eess.AS;audio-visual(comments)
Comments Gen4AVC@ICCV: 1st Workshop on Generative AI for Audio-Visual Content Creation
专题命中 音频语音多模态 :multimodal(abstract,journal_ref)
Journal ref Proceedings of the 2021 International Conference on Multimodal Interaction (ICMI '21), October 18-22, 2021, Montreal, QC, Canada. ACM, New York, NY, USA, 10 pages
专题命中 音频语音多模态 :multimodal(abstract,comments)
Comments Journal on Multimodal User Interface 2021
专题命中 音频语音多模态 :multimodal(abstract,journal_ref)
Journal ref In Proceedings of the 2020 International Conference on Multimodal Interaction, pp. 595-603. 2020
专题命中 音频语音多模态 :分类 cs.CL、cs.AI、eess.AS;multimodal(comments)
Comments Accepted to be published in the First Workshop on Computational Modeling of Human Multimodal Language - ACL 2018
面向推理的后训练与推理时LoRA重缩放:针对音频相关问答任务
专题命中 音频语音多模态 :cross-modal(abstract)
AI总结 该研究针对音频相关问答任务,提出面向推理的LoRA后训练与推理时重缩放方法,在Qwen和MOSS-Audio模型上验证了有效性,提交系统在挑战赛中获总体第三、轻量级系统第二。
Chorus:谐音上下文和传感信号以实现物联网中的无数据模型定制
专题命中 音频语音多模态 :cross-modal(abstract)
AI总结 Chorus通过学习上下文表示,在无需目标域数据的情况下,实现对未知部署条件的模型自适应,实验显示其在多种传感任务中性能优于现有方法,且推理延迟接近传感器部署。