arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Nanyang Technological University(南洋理工大学)

2025-11-25 至 2025-11-25 共收录 12
2511.19425 2025-11-25 cs.CV

SAM3-Adapter: Efficient Adaptation of Segment Anything 3 for Camouflage Object Segmentation, Shadow Detection, and Medical Image Segmentation

SAM3-Adapter: 高效适应Segment Anything 3用于伪装物分割、阴影检测和医学图像分割

Tianrun Chen, Runlong Cao, Xinda Yu, Lanyun Zhu, Chaotao Ding, Deyi Ji, Cheng Chen, Qi Zhu, Chunyan Xu, Papa Mao, Ying Zang

机构 * KOKONI, Moxin (Huzhou) Tech. Co., LTD(摩西(湖州)科技有限公司) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院) School of Information Engineering, Huzhou University(湖州大学信息工程学院) School of Electrical and Electronic Engineering, Nanyang Technological University(新加坡南洋理工大学电子与电气工程学院) College of Computing and Data Science, Nanyang Technological University(新加坡南洋理工大学计算与数据科学学院) School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)

AI总结 SAM3-Adapter通过高效适配框架提升SAM3在伪装物分割、阴影检测和医学影像分割等任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19368 2025-11-25 cs.LG cs.NI

LLM-Driven Stationarity-Aware Expert Demonstrations for Multi-Agent Reinforcement Learning in Mobile Systems

基于大语言模型的站稳意识专家示范的多智能体强化学习在移动系统中的应用

Tianyang Duan, Zongyuan Zhang, Zheng Lin, Songxiao Guo, Xiuxian Guan, Guangyu Wu, Zihan Fang, Haotian Meng, Xia Du, Ji-Zhe Zhou, Heming Cui, Jun Luo, Yue Gao

机构 * Division of Computer Science, The University of Hong Kong(计算机科学系,香港大学) Department of Electrical and Electronic Engineering, The University of Hong Kong(电气电子工程系,香港大学) Department of Computer Science and Technology, Peking University(计算机科学与技术系,北京大学) Department of Computer Science, City University of Hong Kong(计算机科学系,城市大学) China Unicom Digital Technology, China Unicom co.,Ltd(中国联合数字技术,中国联合有限公司) School of Computer and Information Engineering, Xiamen University of Technology(计算机与信息工程学院,厦门理工学院) School of Computer Science, Engineering Research Center of Machine Learning and Industry Intelligence, Sichuan University(计算机科学学院,机器学习与工业智能工程研究中心,四川大学) College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) Institute of Space Internet, Fudan University(空间互联网研究院,复旦大学) School of Computer Science, Fudan University(计算机科学学院,复旦大学)

AI总结 本文提出RELED框架,通过大语言模型驱动的专家示范与自主探索相结合,提升多智能体强化学习在移动系统中的性能和稳定性。

Comments 15 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19168 2025-11-25 cs.LG cs.CL

RAVEN++: Pinpointing Fine-Grained Violations in Advertisement Videos with Active Reinforcement Reasoning

RAVEN++: 通过主动强化推理精准定位广告视频中的细粒度违规

Deyi Ji, Yuekui Yang, Liqun Liu, Peng Shu, Haiyang Wu, Shaogang Tang, Xudong Chen, Shaoping Ma, Tianrun Chen, Lanyun Zhu

机构 * Tencent(腾讯) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Zhejiang University(浙江大学) Nanyang Technological University(南洋理工大学)

AI总结 RAVEN++通过主动强化推理提升广告视频中细粒度违规检测的精度与泛化能力

Comments EMNLP 2025 (Oral, Industry Track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18708 2025-11-25 cs.RO

GVD-TG: Topological Graph based on Fast Hierarchical GVD Sampling for Robot Exploration

GVD-TG:基于快速分层GVD采样的拓扑图用于机器人探索

Yanbin Li, Canran Xiao, Shenghai Yuan, Peilai Yu, Ziruo Li, Zhiguo Zhang, Wenzheng Chi, Wei Zhang

机构 * School of Electronics Engineering, Beijing University of Posts and Telecommunications(北京邮电大学电子工程学院) School of Cyber Science and Technology, Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区计算机科学与技术学院) School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院) Institute for Computer Science, Ludwig Maximilian University of Munich(慕尼黑路德维希-马克西米利安大学计算机科学研究所) Key Lab of Smart Agriculture Systems, China Agricultural University(中国农业大学智能农业系统重点实验室) Robotics and Microsystems Center, School of Mechanical and Electric Engineering, Soochow University(苏州大学机械与电子工程学院机器人与微系统中心)

AI总结 本文提出基于快速分层GVD采样的拓扑图方法,用于提升机器人探索任务中的拓扑地图更新效率与路径规划灵活性。

Comments 12 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13589 2025-11-25 cs.CV

AdaVideoRAG: Omni-Contextual Adaptive Retrieval-Augmented Efficient Long Video Understanding

AdaVideoRAG:多情境自适应检索增强高效长视频理解

Zhucun Xue, Jiangning Zhang, Xurong Xie, Yuxuan Cai, Yong Liu, Xiangtai Li, Dacheng Tao

机构 * Zhejiang University(浙江大学) Youtu Lab, Tencent(腾讯优图实验室) Huazhong University of Science and Technolog(华中科技大学) Nanyang Technological University(南洋理工大学)

AI总结 AdaVideoRAG通过自适应检索增强框架提升长视频理解效率与准确性,支持多层级知识检索与深度语义分析。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20951 2025-11-25 cs.CV

DSOcc: Leveraging Depth Awareness and Semantic Aid to Boost Camera-Based 3D Semantic Occupancy Prediction

DSOcc: 利用深度感知和语义辅助提升基于相机的3D语义占用预测

Naiyu Fang, Zheyuan Zhou, Kang Wang, Ruibo Li, Lemiao Qiu, Shuyou Zhang, Zhe Wang, Guosheng Lin

机构 * S-Lab & College of Computing and Data Science, Nanyang Technological University(S实验室及计算与数据科学学院,南洋理工大学) MMLab, The Chinese University of Hong Kong(M实验室,香港中文大学) State Key Laboratory of Fluid Power & Mechatronic Systems, Zhejiang University(流体动力与机电系统国家重点实验室,浙江大学) SenseTime Research, Hong Kong(商汤研究,香港)

AI总结 DSOcc通过深度感知和语义辅助提升基于相机的3D语义占用预测,实现高效且鲁棒的占用状态和类别推断。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17852 2025-11-25 cs.CV

Sketch-1-to-3: One Single Sketch to 3D Detailed Face Reconstruction

Sketch-1-to-3: 一个单轮廓到三维详细面部重建

Liting Wen, Zimo Yang, Xianlin Zhang, Chi Ding, Mingdao Wang, Xueming Li

机构 * Carnegie Mellon University(卡内基梅隆大学) Nanyang Technological University(南洋理工大学) Beijing University of Posts and Telecommunications(北京邮电大学) Tsinghua University(清华大学)

AI总结 Sketch-1-to-3通过引入GCTD模块和定制损失函数,实现从单个轮廓到高保真三维面部重建的突破性进展。

Comments Accepted by ACM MMAsia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07608 2025-11-25 cs.CV

Faster and Better 3D Splatting via Group Training

更高效且更好的3D点云绘制方法:通过分组训练

Chengbo Wang, Guozheng Ma, Yifei Xue, Yizhen Lao

机构 * Hunan University(湖南大学) Nanyang Technological University(南洋理工大学) Lushan Innovation Lab(庐山创新实验室) Key Laboratory of Digital Culture Smart Design Technology, Minstry of Culture and Tourism(文化与旅游部数字文化智能设计技术重点实验室)

AI总结 本文提出分组训练方法,通过将高斯基本形体分组优化训练效率和渲染质量,实现更高效的3D点云绘制

Comments Accepted to ICCV 2025. Code is available at https://github.com/Chengbo-Wang/3DGS-with-Group-Training

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.03177 2025-11-25 cs.CV

PanoDiffusion: 360-degree Panorama Outpainting via Diffusion

PanoDiffusion:通过扩散模型实现360度全景补全

Tianhao Wu, Chuanxia Zheng, Tat-Jen Cham

机构 * Nanyang Technological University(南洋理工大学) University of Oxford(牛津大学)

AI总结 PanoDiffusion通过双模潜在扩散模型实现360度全景补全,提升全景环绕一致性并生成高质量深度全景图。

Comments Project Page: https://sm0kywu.github.io/panodiffusion/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17943 2025-11-25 cs.CV

SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System

SciEducator: 基于Deming循环多智能体系统的科学视频理解与教育

Zhiyu Xu, Weilong Yan, Yufei Shi, Xin Meng, Tao He, Huiping Zhuang, Ming Li, Hehe Fan

机构 * Jinan University(济南大学) National University of Singapore(新加坡国立大学) Nanyang Technological University(南洋理工大学) Peking University(北京大学) University of Electronic Science and Technology of China(电子科技大学) South China University of Technology(华南理工大学) Guangming Laboratory(光明实验室) Zhejiang University(浙江大学)

AI总结 SciEducator通过Deming循环多智能体系统实现科学视频的自演化理解与教育,生成多模态教学内容并超越现有大语言模型和视频智能体。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17923 2025-11-25 cs.CL cs.AI

Towards Efficient LLM-aware Heterogeneous Graph Learning

面向高效LLM-aware异构图学习

Wenda Li, Tongya Zheng, Shunyu Liu, Yu Wang, Kaixuan Chen, Hanyang Yuan, Bingde Hu, Zujie Ren, Mingli Song, Gang Chen

机构 * State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学) Zhejiang Lab(浙江实验室) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) Zhejiang Provincial Engineering Research Center for Real-Time SmartTech in Urban Security Governance, School of Computer and Computing Science, Hangzhou City University(浙江省实时智能城市安全治理工程技术研究中心,杭州城市大学计算机与计算科学学院) Hangzhou High-Tech Zone (Binjiang) Institute of Blockchain and Data Security, Hangzhou, Zhejiang, China(杭州高新技术区(滨江)区块链与数据安全研究院,杭州,浙江,中国) Nanyang Technological University(南洋理工大学) Bangsun Technology(邦sun科技)

AI总结 本文提出ELLA框架,通过LLM-aware关系分词器和层次关系图变压器,高效解决异构图中复杂关系语义建模和任务间语义间隙问题,实现性能与效率的提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17672 2025-11-25 cs.AI

Cognitive Inception: Agentic Reasoning against Visual Deceptions by Injecting Skepticism

认知 inception:通过注入怀疑进行视觉欺骗的代理推理

Yinjie Zhao, Heng Zhao, Bihan Wen, Joey Tianyi Zhou

机构 * CFAR, Agency for Science, Technology and Research (A*STAR), Singapore(科技研究局(A*STAR)认知智能研究所,新加坡) IHPC, Agency for Science, Technology and Research (A*STAR), Singapore(科技研究局(A*STAR)信息处理中心,新加坡) ROSE Lab, School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院ROSE实验室)

AI总结 本文提出Inception框架,通过注入怀疑提升LLM对视觉欺骗的抵御能力,实现可推广的真伪验证。

详情

展开后加载摘要…

URL PDF HTML 收藏