Journal refZang, X., Jiang, Z., Cheng, J. et al. Instruction-based image editing: a survey on data, models, evaluation, and applications. Vicinagearth 3, 3 (2026)
Symbol and Footprint Database for Electronic Components by Agentic Recognition and Generation
基于智能识别与生成的电子元件符号与引脚封装数据库
Yichen Shi, Yuzhi Liu, Zhuofu Tao, Li Huang, Yuhao Gao, Ting-Jung Lin, Lei Hel
机构
*
Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究院,东方理工大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
BTD Technology(BTD科技)
专题命中
其他VLM
:multimodal large language model(abstract);分类 cs.AI
机构
*
National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(人机混合增强智能国家级重点实验室)
;
Institute of Artificial Intelligence and Robotics(人工智能与机器人研究院)
;
MiLM Plus
;
Xiaomi Inc(小米公司)
;
Zhongguancun Academy(中关村学院)
;
Beijing, China(北京市)
专题命中
其他VLM
:multimodal large language model(abstract);分类 cs.AI
RTE-FM-Dehazer: Radiative Transfer Equation Inspired Flow Matching for Real-World Image Dehazing
RTE-FM-Dehazer: 辐射传输方程启发的流匹配用于真实图像去雾
Chenfeng Wei, Chun Wang, Boyang Zhao, Si Zuo, Shenhong Wang, Chenguang Yang
机构
*
Mashang Consumer Finance Co. Ltd(马上消费金融股份有限公司)
;
Tsinghua University(清华大学)
;
Hunan University(湖南大学)
;
University of Liverpool(利物浦大学)
;
The Hong Kong Polytechnic University(香港理工大学)
Agentic Collaborative Cognition for Zero-Shot 3D Understanding
零样本3D理解的智能体协作认知
Wenxin Wang, Bo Zhang, Feng Chen, Zixuan Wang, Wen Li, Changsheng Li, Yinjie Lei
机构
*
Sichuan University(四川大学)
;
Adelaide University(阿德莱德大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Beijing Institute of Technology(北京理工大学)
专题命中
其他VLM
:multimodal large language model(abstract);分类 cs.CV
NRITYAM: Language Models Meet Art and Heritage of Dance
NRITYAM:语言模型遇见舞蹈的艺术与遗产
Punit Kumar Singh, Niladri Ghosh, Advait Joshiınst, Shailee Choudhary, Michael Färber, Haiqin Yang
机构
*
Shenzhen Technology University(深圳技术大学)
;
New Delhi Institute of Management(新德里管理学院)
;
Technische Universität Dresden(德累斯顿工业大学)
;
Ramakrishna Mission Vivekananda Educational and Research Institute(罗摩克里希纳传道会维韦卡南达教育与研究学院)
;
Indian Institute of Technology(印度理工学院)
;
Swami Vivekananda Institute of Technology(斯瓦米·维韦卡南达技术学院)
;
GuangDong Engineering Technology Research Center of Edge Intelligence(广东省边缘智能工程技术研究中心)
专题命中
其他VLM
:multimodal large language model(abstract);分类 cs.AI
Label Shift Aware Adaptation for Online Zero-shot Learning with Contrastive Language-Image Pre-Training (CLIP)
基于对比语言-图像预训练(CLIP)的在线零样本学习中的标签偏移感知自适应
Pengxiao Han, Changkun Ye, Yanshuo Wang, Jinguang Tong, Miaohua Zhang, Xuesong Li, Jie Hong, Lars Petersson
机构
*
Australian National University(澳大利亚国立大学)
;
China North Vehicle Research Institute(中国北方车辆研究所)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Griffith University(格里菲斯大学)
;
CSIRO(澳大利亚联邦科学与工业研究组织)
;
The University of Hong Kong(香港大学)
机构
*
Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Shenzhen Institute for Advanced Study, University of Electronic Science and Technology of China(电子科技大学深圳高等研究院)
专题命中
其他VLM
:multimodal large language model(abstract);分类 cs.CV
TraRA: Trajectory-level Recognition Aggregation for Video Text Spotting in Urban Surveillance
TraRA: 面向城市监控视频文本识别的轨迹级识别聚合方法
Duc Tri Tran, Trung Thanh Nguyen, Vijay John, Phi Le Nguyen, Yasutomo Kawanishi
机构
*
RIKEN(日本理化学研究所)
;
Hanoi University of Science and Technology(河内科学技术大学)
;
Nagoya University(名古屋大学)
;
Lawrence Technological University(劳伦斯技术大学)
;
Ritsumeikan University(立命馆大学)
VizRAG: Enhancing Retrieval-Augmented Generation with Hypergraph Visualization
VizRAG:通过超图可视化增强检索增强生成
Yanbin Wei, Yang Chen, Renling Gan, Ziru Liu, Xinyu Fu, Chun Kang, Ning Lu, Rui Liu, Yu Zhang, James Kwok
机构
*
Southern University of Science and Technology(南方科技大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Huawei Research(华为研究院)
;
Beihang University(北京航空航天大学)
专题命中
其他VLM
:multimodal large language model(abstract)