WildBox: A Dataset and Benchmark for Aerial Monocular 3D Detection of African Savanna Wildlife
WildBox: 非洲稀树草原野生动物航拍单目3D检测数据集与基准
Vandita Shukla, Kilian Meier, Lucie Laporte-Devylder, Camille Rondeau Saint-Jean, Jenna M. Kline, Blair R. Costelloe, Devis Tuia, Fabio Remondino, Benjamin Risse
机构
*
D Optical Metrology Unit, Fondazione Bruno Kessler(布鲁诺·凯斯勒基金会3D光学计量部)
;
Computer Vision and Machine Learning Systems group, Institute for Geoinformatics, University of Muenster(明斯特大学地理信息学研究所计算机视觉与机器学习系统组)
;
School of Civil, Aerospace and Design Engineering, University of Bristol(布里斯托大学土木、航空航天与设计工程学院)
;
Department of Biology, University of Southern Denmark(南丹麦大学生物学系)
;
The Ohio State University(俄亥俄州立大学)
;
Department of Collective Behavior, Max Planck Institute of Animal Behavior(马克斯·普朗克动物行为研究所集体行为系)
;
University of Konstanz(康斯坦茨大学)
;
Centre for the Advanced Study of Collective Behaviour, University of Konstanz(康斯坦茨大学集体行为高级研究中心)
;
Department of Biology, University of Konstanz(康斯坦茨大学生物学系)
;
Environmental Computational Science and Earth Observation Laboratory, EPFL(瑞士联邦理工学院环境计算科学与地球观测实验室)
MODE-RAG: Manifold Outlier Diagnosis and Energy-based Retrieval-Augmented Generation Evaluation
MODE-RAG: 基于流形异常诊断和能量的检索增强生成评估
Zehang Wei, Jiaxin Dai, Jiamin Yan, Xiang Xiang
机构
*
School of Computer Science & Tech, Huazhong University of Science and Technology(华中科技大学计算机科学与技术学院)
;
School of AI and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院)
GRACE: Boosting Video MLLMs with Grounded Action-Centric Evidence for Viewer Sentiment Prediction
GRACE: 基于接地动作中心证据增强视频多模态大语言模型用于观众情感预测
Ruoxuan Yang, Tieyuan Chen, Xiaofeng Huang, Haibing Yin, Jun Wang, Xiping Chen, Jun Yin, Xuesong Gao, Weiyao Lin
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Hangzhou Dianzi University(杭州电子科技大学)
;
The 52nd Research Institute of China Electronics Technology Group Corporation(中国电子科技集团公司第五十二研究所)
;
Hangzhou Bywin Technology Co., Ltd.(杭州百威科技有限公司)
;
Zhejiang Dahua Technology Co., Ltd.(浙江大华技术股份有限公司)
;
School of Information Science and Engineering, Shandong University(山东大学信息科学与工程学院)
;
Haihe Laboratory of Information Technology Application Innovation(海河信息技术应用创新实验室)
专题命中
评测与基准
:large language model(abstract);language model(abstract)
Comments10 pages, 3 figures, 2 Tables, conference KES-2026 30th International Conference on Knowledge-Based and Intelligent Information & Engineering Systems
机构
*
Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Shenzhen Institute for Advanced Study, University of Electronic Science and Technology of China(电子科技大学深圳高等研究院)
专题命中
评测与基准
:large language model(abstract);language model(abstract)
TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Improved Vision-Language Alignment
TEVI: 基于稀疏自编码器的文本条件视觉表示编辑以改进视觉-语言对齐
Sweta Mahajan, Sukrut Rao, Jiahao Xie, Alexander Koller, Bernt Schiele
机构
*
Max Planck Institute for Informatics, Saarland Informatics Campus, Saarbrücken, Germany(马克斯·普朗克研究所信息学院,萨尔兰信息学院,德国萨尔布吕肯)
;
Department of Language Science and Technology, Saarland University, Saarbrücken, Germany(语言科学与技术系,萨尔兰大学,德国萨尔布吕肯)