arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 46430 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 音频语音多模态 4597 篇

2307.05451 2023-07-12 cs.HC 50%

Detection Threshold of Audio Haptic Asynchrony in a Driving Context

Gyanendra Sharma, Hiroshi Yasuda, Manuel Kuehner

专题命中 音频语音多模态 :multimodal(abstract)

Comments 8 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.00124 2023-06-22 cs.LO 50%

Axiomatizing Hybrid XPath with Data

Carlos Areces, Raul Fervari

专题命中 音频语音多模态 :multi-modal(abstract)

Journal ref Logical Methods in Computer Science, Volume 17, Issue 3 (July 20, 2021) lmcs:6259

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.11600 2023-06-16 cs.SI cs.CR 50%

Creative beyond TikToks: Investigating Adolescents' Social Privacy Management on TikTok

Nico Ebert, Tim Geppert, Joanna Strycharz, Melanie Knieps, Michael Hönig, Elke Brucker-Kley

专题命中 音频语音多模态 :audio-visual(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.17180 2023-05-30 cs.HC 50%

Exploring Human Response Times to Combinations of Audio, Haptic, and Visual Stimuli from a Mobile Device

Kyle T. Yoshida, Joel X. Kiernan, Allison M. Okamura, Cara M. Nunez

专题命中 音频语音多模态 :multi-modal(abstract)

Comments Accepted to World Haptics Conference 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.17067 2023-05-29 cs.DL 50%

Cluster Analysis of Open Research Data and a Case for Replication Metadata

Ana Trisovic

专题命中 音频语音多模态 :audio-visual(abstract)

Journal ref 2022 IEEE 18th International Conference on e-Science (e-Science), Salt Lake City, UT, USA, 2022, pp. 423-424

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.03789 2023-05-17 cs.CY 50%

Broadening AI Ethics Narratives: An Indic Art View

Ajay Divakaran, Aparna Sridhar, Ramya Srinivasan

专题命中 音频语音多模态 :multimodal(abstract)

Journal ref 2023 ACM Conference on Fairness, Accountability, and Transparency

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.05635 2023-05-10 astro-ph.IM 50%

Research on access, use and effective exploration of astronomical observational and bibliographical data from sonification

Johanna Casado, Beatriz García

专题命中 音频语音多模态 :multimodal(abstract)

Comments Thesis of 265 pages with bibliography, in Spanish language, the rest are appends

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.00884 2023-03-03 cs.HC 50%

Encouraging Emotion Regulation in Social Media Conversations through Self-Reflection

Akriti Verma, Shama Islam, Valeh Moghaddam, Adnan Anwar

专题命中 音频语音多模态 :multimodal(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.14717 2023-03-01 cs.HC 50%

Bridging the Generational Gap: Exploring How Virtual Reality Supports Remote Communication Between Grandparents and Grandchildren

Xiaoying Wei, Yizheng Gu, Emily Kuang, Xian Wang, Beiyan Cao, Xiaofu Jin, Mingming Fan

专题命中 音频语音多模态 :audio-visual(abstract)

Comments Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (CHI '23), April 23--28, 2023, Hamburg, Germany

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.03088 2023-02-08 cs.RO cs.HC 50%

Sketching Robot Programs On the Fly

David Porfirio, Laura Stegner, Maya Cakmak, Allison Sauppé, Aws Albarghouthi, Bilge Mutlu

专题命中 音频语音多模态 :multimodal(abstract)

Comments Accepted at HRI '23, March 13-16, 2023, Stockholm, Sweden

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.03537 2023-01-24 cs.AR 50%

TinyVers: A Tiny Versatile System-on-chip with State-Retentive eMRAM for ML Inference at the Extreme Edge

Vikram Jain, Sebastian Giraldo, Jaro De Roose, Linyan Mei, Bert Boons, Marian Verhelst

专题命中 音频语音多模态 :multi-modal(abstract)

Comments Accepted in IEEE Journal of Solid-State Circuits

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.00964 2023-01-04 cs.RO cs.HC cs.LG 50%

e-Inu: Simulating A Quadruped Robot With Emotional Sentience

Abhiruph Chakravarty, Jatin Karthik Tripathy, Sibi Chakkaravarthy S, Aswani Kumar Cherukuri, S. Anitha, Firuz Kamalov, Annapurna Jonnalagadda

专题命中 音频语音多模态 :audio-visual(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.08849 2022-11-22 cs.HC 50%

Affective visualization in Virtual Reality: An integrative review

Andres Pinilla, Jaime Garcia, William Raffe, Jan-Niklas Voigt-Antons, Robert Spang, Sebastian Möller

专题命中 音频语音多模态 :audio-visual(abstract)

Journal ref Frontiers in Virtual Reality, 2, 630731 (2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.02753 2022-11-21 cs.DB cs.LG 50%

The Tensor Data Platform: Towards an AI-centric Database System

Apurva Gandhi, Yuki Asada, Victor Fu, Advitya Gemawat, Lihao Zhang, Rathijit Sen, Carlo Curino, Jesús Camacho-Rodríguez, Matteo Interlandi

专题命中 音频语音多模态 :multi-modal(abstract)

Comments Accepted for publication at The Conference on Innovative Data Systems Research (CIDR) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.06943 2022-11-08 cs.IT eess.SP math.IT 50%

Understand-Before-Talk (UBT): A Semantic Communication Approach to 6G Networks

Shiva Raj Pokhrel, Jinho Choi

专题命中 音频语音多模态 :multi-modal(abstract)

Journal ref IEEE Transactions on Vehicular Technology 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.12381 2022-10-04 cs.LG math.PR math.ST stat.ML stat.TH 50%

Convergence of score-based generative modeling for general data distributions

Holden Lee, Jianfeng Lu, Yixin Tan

专题命中 音频语音多模态 :multimodal(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.09580 2022-08-23 cs.RO cs.HC 50%

Using Affect as a Communication Modality to Improve Human-Robot Communication in Robot-Assisted Search and Rescue Scenarios

Sami Alperen Akgun, Moojan Ghafurian, Mark Crowley, Kerstin Dautenhahn

专题命中 音频语音多模态 :multi-modal(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.10828 2022-07-25 cs.HC 50%

Tell Me How You Feel: Designing Emotion-Aware Voicebots to Ease Pandemic Anxiety In Aging Citizens

W. Mieleszczenko-Kowszewicz, K. Warpechowski, K. Zieliński, R. Nielek, A. Wierzbicki

专题命中 音频语音多模态 :multimodal(abstract)

Comments 16 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.04344 2022-07-12 cs.HC 50%

Freedom to Choose: Understanding Input Modality Preferences of People with Upper-body Motor Impairments for Activities of Daily Living

Franklin Mingzhe Li, Michael Xieyang Liu, Yang Zhang, Patrick Carrington

专题命中 音频语音多模态 :multimodal(abstract)

Comments ASSETS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.10967 2022-06-23 cs.HC 50%

Audience Response Prediction from Textual Context

Ibrahim Shoer, Berker Turker, Engin Erzin

专题命中 音频语音多模态 :audio-visual(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.04461 2022-06-17 cs.HC 50%

Virtual Reality for Emotion Elicitation -- A Review

Rukshani Somarathna, Tomasz Bednarz, Gelareh Mohammadi

专题命中 音频语音多模态 :audio-visual(abstract)

Comments This work is submitted to IEEE for possible publication

Journal ref IEEE Transactions on Affective Computing, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.12494 2022-05-20 cs.LG cs.RO stat.ML 50%

SEMI: Self-supervised Exploration via Multisensory Incongruity

Jianren Wang, Ziwen Zhuang, Hang Zhao

专题命中 音频语音多模态 :audio-visual(abstract)

Comments Accepted at ICRA 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.07140 2022-04-15 eess.SP cs.LG 50%

The Pseudo Projection Operator: Applications of Deep Learning to Projection Based Filtering in Non-Trivial Frequency Regimes

Matthew L. Weiss, Nathan C. Frey, Siddharth Samsi, Randy C. Paffenroth, Vijay Gadepally

专题命中 音频语音多模态 :multi-modal(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.11487 2022-02-22 cs.LG 50%

Routine Clustering of Mobile Sensor Data Facilitates Psychotic Relapse Prediction in Schizophrenia Patients

Joanne Zhou, Bishal Lamichhane, Dror Ben-Zeev, Andrew Campbell, Akane Sano

专题命中 音频语音多模态 :multimodal(abstract)

Comments JMIR mHealth and uHealth

详情

展开后加载摘要…

URL PDF HTML 收藏
1502.02796 2022-02-09 cs.HC cs.CY 50%

Opportunistic and Context-aware Affect Sensing on Smartphones: The Concept, Challenges and Opportunities

Rajib Rana, Margee Hume, John Reilly, Raja Jurdak, Jeffrey Soar

专题命中 音频语音多模态 :audio-visual(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.11167 2022-01-28 cs.HC cs.RO 50%

Artificial Emotional Intelligence in Socially Assistive Robots for Older Adults: A Pilot Study

Hojjat Abdollahi, Mohammad H. Mahoor, Rohola Zandie, Jarid Siewierski, Sara H. Qualls

专题命中 音频语音多模态 :multimodal(abstract)

Comments To be published in IEEE Transactions on Affective Computing

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.09872 2021-12-03 cs.LG cs.CR 50%

Robust Sensor Fusion Algorithms Against Voice Command Attacks in Autonomous Vehicles

Jiwei Guan, Xi Zheng, Chen Wang, Yipeng Zhou, Alireza Jolfa

专题命中 音频语音多模态 :multimodal(abstract)

Comments 8 pages, 2 tables, 9 figures

Journal ref IEEE Trustcom 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.07851 2021-11-16 cs.HC 50%

Park4U Mate: Context-Aware Digital Assistant for Personalized Autonomous Parking

Antonyo Musabini, Evin Bozbayir, Hervé Marcasuzaa, Omar Adair Islas Ramírez

专题命中 音频语音多模态 :multi-modal(abstract)

Comments Accepted at 2021 IEEE Intelligent Vehicles Symposium - IV (matching camera-ready version)

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.06288 2021-11-12 cs.HC 50%

MaTIC: a Mathematical Theory of Inferential Communication

Joan Llobera, Jordi Vallverdú

专题命中 音频语音多模态 :multi-modal(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.01237 2021-11-03 physics.med-ph 50%

Auditory-visual scenes for hearing research

Steven van de Par, Stephan D. Ewert, Lubos Hladek, Christoph Kirsch, Julia Schütze, Josep Llorca-Bofí, Giso Grimm, Maartje M. E. Hendrikse, Birger Kollmeier, Bernhard U. Seeber

专题命中 音频语音多模态 :audio-visual(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏