VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents
Xiao Liu, Tianjie Zhang, Yu Gu, Iat Long Iong, Xixuan Song +16 more
2/12/2026
No confident code match yet
We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 4 weaker matches are hidden by default — verify before relying on them.
No code implementations found yet.
Know of an implementation? Let us know in the comments below!
@article{liu2026visualagentbench,
title = {VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents},
author = {Xiao Liu and Tianjie Zhang and Yu Gu and Iat Long Iong and Xixuan Song and Yifan Xu and Shudan Zhang and Hanyu Lai and Jiadai Sun and Xinyue Yang and Yu Yang and Zehan Qi and Shuntian Yao and Xueqiao Sun and Siyi Cheng and Qinkai Zheng and Hao Yu and Hanchen Zhang and Wenyi Hong and Ming Ding and Chan Hee Song},
year = {2026},
url = {https://api.semanticscholar.org/CorpusID:e47571dddffa3256f534f9eb5c606ffe64051352},
journal = {ICLR 2025 2025}
}