Exploring Automated Recognition of Instructional Activity and Discourse from Multimodal Classroom Data
探索多模态课堂数据中教学活动和话语的自动化识别
Ivo Bueno, Ruikun Hou, Babette Bühler, Tim Fütterer, James Drimalla, Jonathan Kyle Foster, Peter Youngs, Peter Gerjets, Ulrich Trautwein, Enkelejda Kasneci
CAPE: A CLIP-Aware Pointing Ensemble of Complementary Heatmap Cues for Embodied Reference Understanding
CAPE:一种基于CLIP的互补热图线索点集用于具身参照理解
Fevziye Irem Eyiokur, Dogucan Yaman, Hazım Kemal Ekenel, Alexander Waibel
机构
*
Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)
;
Istanbul Technical University(伊斯坦布尔技术大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
KIT Campus Transfer GmbH (KCT)(KIT校园转移有限责任公司)
Multi-Modal Graph Convolutional Network with Sinusoidal Encoding for Robust Human Action Segmentation
多模态图卷积网络与正弦编码用于鲁棒的人体动作分割
Hao Xing, Kai Zhe Boey, Yuankai Wu, Darius Burschka, Gordon Cheng
机构
*
Institute for Cognitive Systems(认知系统研究所)
;
Chair of Media Technology(媒体技术教授职位)
;
Machine Vision and Perception Group(机器视觉与感知小组)
;
School of Computation, Information and Technology(计算、信息与技术学院)
RAD: Towards Trustworthy Retrieval-Augmented Multi-modal Clinical Diagnosis
RAD:迈向可信的检索增强多模态临床诊断
Haolin Li, Tianjie Dai, Zhe Chen, Siyuan Du, Jiangchao Yao, Ya Zhang, Yanfeng Wang
机构
*
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
CMIC, Shanghai Jiao Tong University(上海交通大学计算机学院)
;
School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院)
;
Institute of Artificial Intelligence for Medicine, Shanghai Jiao Tong University(上海交通大学医学人工智能研究所)
Beyond Pixels: A Training-Free, Text-to-Text Framework for Remote Sensing Image Retrieval
超越像素:一种无训练的文本到文本框架用于遥感图像检索
J. Xiao, Y. Guo, X. Zi, K. Thiyagarajan, C. Moreira, M. Prasad
机构
*
Information Technology University of Technology Sydney Sydney, Australia(信息科技技术大学悉尼分校悉尼澳大利亚)
;
Robotics Laboratory (SensR Lab) Centre for Advanced Manufacturing Technology Western Sydney University Sydney, Australia(机器人实验室(SensR实验室)先进制造技术中心西悉尼大学悉尼澳大利亚)
;
The Data Science Institute Faculty of Engineering(数据科学学院工程学院)
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Stanford University(斯坦福大学)
;
Nanyang Technological University(南洋理工大学)
;
SenseTime Group Ltd(商汤科技有限公司)
;
Beihang University(北京航空航天大学)
;
Nokia Bell Labs(诺基亚贝尔实验室)
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
SAVE:基于稀疏自编码器的视觉信息增强用于缓解物体幻觉
Sangha Park, Seungryong Yoo, Jisoo Mok, Sungroh Yoon
机构
*
Department of Electrical and Computer Engineering, Seoul National University(电子与计算机工程系,首尔国立大学)
;
Daegu Gyeongbuk Institute of Science and Technology(大邱庆州科学技术院)
;
IPAI, AIIS, ASRI, INMC, and ISRC, Seoul National University(IPAI、AIIS、ASRI、INMC 和 ISRC,首尔国立大学)
机构
*
State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Artificial Intelligence Research Institute(人工智能研究院)
;
iFLYTEK Co., Ltd(iFLYTEK公司)
Comments6 pages, 2 figures, 1 table. Accepted at UCC '25 (IEEE/ACM 18th International Conference on Utility and Cloud Computing), December 1-4, 2025, Nantes, France. DOI to be activated upon final publication
Salient Object Detection in Complex Weather Conditions via Noise Indicators
通过噪声指示器进行复杂天气条件下的显著物体检测
Quan Chen, Xiaokai Yang, Tingyu Wang, Rongfeng Lu, Xichun Sheng, Yaoqi Sun, Chenggang Yan
机构
*
School of Automation, Hangzhou Dianzi University(自动化学院,杭州电子大学)
;
College of Artificial Intelligence, Jiaxing University(人工智能学院,嘉兴大学)
;
School of Communication Engineering, Hangzhou Dianzi University(通信工程学院,杭州电子大学)
;
School of Faculty of Applied Science, Macao Polytechnic University(应用科学学院,澳门理工学院)
;
School of Mathematics and Computer Science, Lishui University(数学与计算机科学学院,丽水大学)
;
Lishui Institute of Hangzhou Dianzi University(杭州电子大学丽水学院)