MERIT: Memory-Enhanced Retrieval for Interpretable Knowledge Tracing
MERIT: 用于可解释知识追踪的记忆增强检索
Runze Li, Kedi Chen, Guwei Feng, Mo Yu, Jun Wang, Wei Zhang
机构
*
School of Computer Science and Technology, East China Normal University(东华师范大学计算机科学与技术学院)
;
Shanghai Innovation Institute(上海创新研究院)
;
WeChat AI, Tencent(微信AI,腾讯)
3DSceneEditor: Controllable 3D Scene Editing with Gaussian Splatting
3DSceneEditor: 基于高斯散射的可控3D场景编辑
Ziyang Yan, Yihua Shao, Minwen Liao, Siyu Chen, Nan Wang, Muyuan Lin, Jenq-Neng Hwang, Hao Zhao, Fabio Remondino, Lei Li
机构
*
D Optical Metrology unit, Fondazione Bruno Kessler(3D光学测绘单元,布鲁诺·凯斯勒基金会)
;
Department of Information Engineering and Computer Science, University of Trento(信息工程与计算机科学系,特伦托大学)
Mitigating Objectness Bias and Region-to-Text Misalignment for Open-Vocabulary Panoptic Segmentation
缓解对象性偏差和区域到文本对齐问题以实现开放词汇全景分割
Nikolay Kormushev, Josip Šarić, Matej Kristan
机构
*
University of Ljubljana(卢布尔雅那大学)
;
ETH Zurich(苏黎世联邦理工学院)
;
University of Zagreb(扎格reb大学)
;
Faculty of Comp. and Inf. Science(计算机与信息科学系)
;
Dept. of Computer Science(计算机科学系)
;
Faculty of Elec. Eng. and Computing(电子工程与计算科学系)
NoOVD: Novel Category Discovery and Embedding for Open-Vocabulary Object Detection
NoOVD:开放词汇物体检测中的新类别发现与嵌入
Yupeng Zhang, Ruize Han, Zhiwei Chen, Wei Feng, Liang Wan
机构
*
College of Intelligence and Computing, Tianjin University(天津大学智能与计算学院)
;
Key Research Center for Surface Monitoring and Analysis of Relics, State Administration of Cultural Heritage(文物表面监测与分析国家重点研究中心)
;
Faculty of Computer Science and Artificial Intelligence, Shenzhen University of Advanced Technology(深圳先进技术大学计算机科学与人工智能学院)
;
School of Artificial Intelligence, Nanchang University(南昌大学人工智能学院)
机构
*
Department of Management Science and Technology(管理科学与技术系)
;
Hellenic Mediterranean University(希伯伦地中海大学)
;
Institute of Computer Science, FORTH(信息科学研究所,FORTH)
GenVideoLens: Where LVLMs Fall Short in AI-Generated Video Detection?
GenVideoLens:在AI生成视频检测中LVLMs的局限性
Yueying Zou, Pei Pei Li, Zekun Li, Xinyu Guo, Xing Cui, Huaibo Huang, Ran He
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
University of California, Santa Barbara(加州大学圣巴巴拉分校)
;
Center for Research on Intelligent Perception and Computing, NLPR, Institute of Automation, Chinese Academy of Sciences(智能感知与计算中心、国家智能感知与信息处理实验室、中国科学院自动化研究所)
Evidence Packing for Cross-Domain Image Deepfake Detection with LVLMs
跨领域图像深度伪造检测中的证据打包与大视觉语言模型
Yuxin Liu, Fei Wang, Kun Li, Yiqi Nie, Junjie Chen, Zhangling Duan, Zhaohong Jia
机构
*
Anhui University(安徽大学)
;
Hefei University of Technology(合肥工业大学)
;
IAI, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院)
;
United Arab Emirates University(阿拉伯联合酋长国大学)
Concept-to-Pixel: Prompt-Free Universal Medical Image Segmentation
概念到像素:无提示通用医学图像分割
Haoyun Chen, Fenghe Tang, Wenxin Ma, Shaohua Kevin Zhou
机构
*
School of Biomedical Engineering, Division of Life Sciences
;
Medicine, University of Science
;
Technology of China (USTC), Hefei, Anhui 230026, China Center for Medical Imaging, Robotics, Analytic Computing \& Learning (MIRACLE), Suzhou Institute for Advanced Research, USTC, Suzhou, Jiangsu 215123, China Jiangsu Provincial Key Laboratory of Multimodal Digital Twin Technology, Suzhou Jiangsu, 215123, China State Key Laboratory of Precision
专题命中
视觉定位与Grounding
:multimodal large language model(abstract);分类 cs.CV
Deploying Semantic ID-based Generative Retrieval for Large-Scale Podcast Discovery at Spotify
在Spotify上部署基于语义ID的生成检索用于大规模播客发现
Edoardo D'Amico, Marco De Nadai, Praveen Chandar, Divita Vohra, Shawn Lin, Max Lefarov, Paul Gigioli, Gustavo Penha, Ilya Kopysitsky, Ivo Joel Senese, Darren Mei, Francesco Fabbri, Oguz Semerci, Yu Zhao, Vincent Tang, Brian St. Thomas, Alexandra Ranieri, Matthew N. K. Smith, Aaron Bernkopf, Bryan Leung, Ghazal Fazelnia, Mark VanMiddlesworth, Timothy Christopher Heath, Petter Pehrson Skiden, Alice Y. Wang, Doug J. Cole, Andreas Damianou, Maya Hristakeva, Reid Wilbur, Tarun Chillara, Vladan Radosavljevic, Pooja Chitkara, Sainath Adapa, Juan Elenter, Bernd Huber, Jacqueline Wood, Saaketh Vedantam, Jan Stypka, Sandeep Ghael, Martin D. Gould, David Murgatroyd, Yves Raimond, Mounia Lalmas, Paul N. Bennett
PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part Segmentation
PCA-Seg:重新审视开放词汇语义和部分分割中的成本聚合
Jianjian Yin, Tao Chen, Yi Chen, Gensheng Pei, Xiangbo Shu, Yazhou Yao, Fumin Shen
机构
*
Nanjing University of Science and Technology(南京理工大学)
;
Nanjing Normal University(南京师范大学)
;
Department of Electrical and Computer Engineering, Sungkyunkwan University(成均馆大学电子与计算机工程系)
;
University of Electronic Science and Technology of China(电子科技大学)
;
State Key Laboratory of Intelligent Manufacturing of Advanced Construction Machinery(先进施工机械智能制造国家重点实验室)