arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

2026-08-26 至 2026-08-26 共收录 14
2608.24535 2026-08-26 cs.CV cs.HC 新提交

VizAnchor: Decoding Manipulation Intent from Tampering Visualizations via Dual-Anchor Reasoning

VizAnchor:通过双锚推理从篡改可视化中解码操纵意图

Xiaotian Zhang, Huayuan Ye, Haiyang Zhang, Chenhui Li, Changbo Wang, Sicheng Song

机构 * School of Data Science and Engineering, East China Normal University(华东师范大学数据科学与工程学院) Division of Emerging Interdisciplinary Areas, The Hong Kong University of Science and Technology(香港科技大学新兴跨学科领域分部) School of Computer Science and Technology, East China Normal University(华东师范大学计算机科学与技术学院)

AI总结 本文提出VizAnchor框架,通过双锚构建与VLM推理实现可视化操纵理解,含三个专用智能体,构建相关数据集,可准确定位篡改并生成忠实解释。

Comments 39 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.24441 2026-08-26 cs.AI 新提交

A Behavior-Guided Online Probabilistic Forecasting Method for Electric vehicle Charging Loads

面向电动汽车充电负荷的行为引导型在线概率预测方法

Chenghan Li, Qingxiang Liu, Yinliang Xu, Yuxuan Liang

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) Intelligent Transportation Thrust, The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)智能交通学域)

AI总结 针对电动汽车充电负荷的行为异质性与时间变异性挑战,提出行为引导型在线概率预测框架,经10个真实充电站实验,其在1小时、4小时预测中均显著优于基线方法,实现稳定性能提升。

Comments Submitted to IEEE

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.24439 2026-08-26 cs.CV 新提交

DoublesEval: Diagnosing Multi-Agent Tactical Reasoning in Vision-Language Models via Professional Doubles Badminton

DoublesEval:通过专业双打羽毛球诊断视觉语言模型的多智能体战术推理能力

Jintao Cheng, Weibin Li

机构 * The Hong Kong University of Science and Technology(香港科技大学) University of Macau(澳门大学)

AI总结 本研究提出DoublesEval框架,以专业双打羽毛球为测试平台评估VLMs的多智能体战术推理能力,发现现有模型存在明显瓶颈,所提TacticCheck可提升模型表现但仍有差距,强调需改进VLMs的评估范式。

Comments Accepted by BMVC2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.24372 2026-08-26 cs.CV 新提交

Bridging Adversarial and Collaborative Learning for AI-Generated Image Quality Assessment

桥接对抗学习与协同学习的AI生成图像质量评估

Baoliang Chen, Qing Lin, Sijie Mai

机构 * South China Normal University(华南师范大学) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 该研究针对AI生成图像质量评估中感知保真度与提示对齐度的交互问题,提出含对抗与协同推理通路的交互感知框架,在基准测试中达最优准确率且具可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.24086 2026-08-26 cs.AI cs.CE cs.SE 新提交

EMRB: A Multi-Level Benchmark for Evaluating LLM Reasoning over Raw Electromagnetic Signals

EMRB:用于评估大型语言模型对原始电磁信号推理能力的多级基准

Mingxu Zhang, Ying Sun, Yuhan Li, Yang Ji, Dazhong Shen, Ke Zhang, Shan Huang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) The 63rd Research Institute, National University of Defense Technology(国防科技大学第六十三研究所) Nanjing University of Aeronautics and Astronautics(南京航空航天大学)

AI总结 本研究提出EMRB基准评估LLMs对原始电磁信号的推理能力,针对14个LLMs的实验显示其表现存在差距,提出的ReconPilot方法可显著提升LLMs的相关任务得分。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.23720 2026-08-26 cs.CV 新提交

Platonic Representation Hypothesis on World Models

世界模型的柏拉图式表征假说

Wenhow Li (1), Chengwei MA (1), Hui Xiong (1), Ying-Cong Chen (1), Lei Zhang (1) ((1) The Hong Kong University of Science and Technology (Guangzhou), Guangzhou, China)

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 该研究针对世界模型的表征性质,提出预测一致性假设,通过DINO-WM实验发现性能良好的世界模型会形成几何相似的内部结构,且模型特征可跨模型映射,证实预测一致性能促进共享潜在结构的形成。

Comments 18 pages, 10 figures, 2 tables. Wenhow Li and Chengwei MA contributed equally. Project page: this https URL (https://sellerbubble.github.io/platonic-representation-hypothesis-on-world-models/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.23719 2026-08-26 cs.CL 新提交

ADE: Agentic Data Evolution Framework for Human-Centered Objectives

ADE:面向以人为中心目标的智能体数据演化框架

Yang Yu, Yilin Jiang, Zexuan Fei, Yiming Luo, Xingkai Song, Kaiyi Huang, Aimin Zhou, Xin Lin, Fei Tan

机构 * East China Normal University(华东师范大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Shanghai Innovation Institute(上海创新研究院)

AI总结 针对大语言模型与以人为中心目标对齐的难题,提出以数据为中心的ADE框架,通过OVS流程与稳态准入机制提升数据质量,在多基准及不同场景下均实现性能提升。

Comments accepted by EMNLP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.22888 2026-08-26 cs.CV 版本更新

NemoSplat: Feed-Forward 4D Gaussian Splatting for Media-Aware Underwater Reconstruction

NemoSplat:面向介质感知水下重建的前馈式4D高斯溅射方法

Xiaopeng Guo, Wai Chung Tse, Yipeng Zhu, Hanwen Zhang, Huajian Huang, Sai-Kit Yeung

机构 * The Hong Kong University of Science and Technology(香港科技大学) Beijing Institute of Technology(北京理工大学)

AI总结 针对水下环境光散射与动态物体导致的重建难题,提出首个介质感知前馈4D高斯溅射框架NemoSplat,设计可提示动态解缠器与介质感知高斯预测器,构建大规模水下数据集,实现了最优跟踪精度与高保真渲染。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.07209 2026-08-26 cs.CR cs.LG 版本更新

Continual Learning With Participation Privacy: An Auditable Buffering-Aggregation Recipe

具有参与隐私的持续学习:一种可审计的缓冲聚合方法

T-H. Hubert Chan, Elaine Shi, Mengshi Zhao, Mingxun Zhou

机构 * The University of Hong Kong(香港大学) Carnegie Mellon University(卡内基梅隆大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

AI总结 研究单编辑相邻用户流的持续学习隐私问题,提出模块化方法,用随机缓冲包装器处理,证明认证定理,使标准原语产生轨迹级差分隐私,建立隐私与延迟联系。

Comments This version corrects and clarifies the independent-decomposability condition underlying the adaptive-safety result in the pre-conference version of the ICML 2026 paper, with corresponding revisions to the affected statements and proofs

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.11263 2026-08-26 math.ST cs.LG math.NA math.PR 版本更新

Geometric bias in eigenspace perturbation under random heterogeneous noise

随机异质噪声下特征空间扰动的几何偏差

Fengkai Liu, Ke Wang, Wanjie Wang

机构 * Department of Mathematics, Hong Kong University of Science and Technology(香港科技大学数学系) Department of Statistics and Data Science, National University of Singapore(新加坡国立大学统计与数据科学系)

AI总结 针对稀疏、异质方差噪声下的信号加噪声矩阵,研究发现经验特征向量存在经典扰动界无法捕捉的系统性几何偏差,并通过二次向量方程和精细各向同性局部律推导了最优非渐近扰动界。

Comments 132 pages, 2 figures. Restructured the paper and added new applications of the main perturbation results

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.26757 2026-08-26 cs.RO 版本更新

Beyond Viewpoint Generalization: What Multi-View Demonstrations Offer and How to Synthesize Them for Robot Manipulation?

超越视角泛化:多视角演示提供了什么以及如何为机器人操作合成它们

Boyang Cai, Qiwei Liang, Jiawei Li, Shihang Weng, Zhaoxin Zhang, Tao Lin, Xiangyu Chen, Wenjie Zhang, Jiaqi Mao, Weisheng Xu, Bin Yang, Jiaming Liang, Junhao Cai, Renjing Xu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Shenzhen University(深圳大学) Beijing Jiaotong University(北京交通大学) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))

AI总结 研究通过实验验证多视角演示在机器人操作中的性能提升,揭示了多视角数据在提升泛化能力和突破单视角限制方面的优势,提出RoboNVS框架合成多视角数据以提升下游政策表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10485 2026-08-26 cs.AI 版本更新

Panning for Gold: Expanding Domain-Specific Knowledge Graphs with General Knowledge

寻找黄金:利用通用知识图谱扩展领域特定知识图谱

Runhao Zhao, Weixin Zeng, Wentao Zhang, Chong Chen, Zhengpin Li, Xiang Zhao, Lei Chen

机构 * National Key Laboratory of Big Data and Decision(大数据与决策国家重点实验室) National University of Defense Technology(国防科技大学) The Center for machine learning research(机器学习研究中心) Peking University(北京大学) Tsinghua University(清华大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

AI总结 本文提出领域特定知识图谱融合任务,通过神经符号框架ExeFuse,利用通用知识图谱提升领域知识图谱的完整性和实用性。

Comments 13 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11239 2026-08-26 cs.LG 版本更新

Massive-STEPS: Massive Semantic Trajectories for Understanding POI Check-ins -- Dataset and Benchmarks

Massive-STEPS: 用于理解POI签到的大量语义轨迹 -- 数据集和基准测试

Wilson Wongso, Hao Xue, Flora D. Salim

机构 * University of New South Wales(新南威尔士大学) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 Massive-STEPS通过提供大规模、公开的POI签到数据集和基准测试,推动人类移动性和POI轨迹建模的可重复研究。

Comments Accepted to SIGSPATIAL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07737 2026-08-26 cs.CV cs.AI 版本更新

Seeing vs. Believing: Evaluating the Language Bias of Open-Source MLLMs in Counter-Intuitive Scenes

看见 vs. 相信:评估开源多模态大模型在反直觉场景中的语言偏见

Chen Ling, Tongwei Zhang, Hanqian Li, Nai Ding

机构 * Zhejiang University(浙江大学) Beijing University of Posts and Telecommunications(北京邮电大学) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 为评估多模态大模型处理反直觉动作场景的能力,提出CAIT基准(400个高保真合成场景),发现开源模型因语言先验而忽视视觉证据,性能接近随机水平,而链式思维推理虽提升准确率但导致过度思考拒绝视觉内容,通过微调和结构化提示可缓解此偏见。

详情

展开后加载摘要…

URL PDF HTML 收藏