Visual Chronicles: Using Multimodal LLMs to Analyze Massive Collections of Images
机构 * Stanford University(斯坦福大学) ; Google DeepMind(谷歌DeepMind)
专题命中 其他VLM :MLLM(abstract);分类 cs.CV、cs.AI
Comments ICCV 2025, Project page: https://boyangdeng.com/visual-chronicles , second and third listed authors have equal contributions