Video Token Merging for Long Video Understanding
Seon-Ho Lee, Jue Wang, Zhikang Zhang, D. Fan, Xinyu Li
2/3/2026
No confident code match yet
We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 1 weaker match is hidden by default — verify before relying on them.
No code implementations found yet.
Know of an implementation? Let us know in the comments below!
@article{lee2026video,
title = {Video Token Merging for Long Video Understanding},
author = {Seon-Ho Lee and Jue Wang and Zhikang Zhang and D. Fan and Xinyu Li},
year = {2026},
url = {https://www.semanticscholar.org/paper/cd7c76dccf476fda819b7df4cc53044cf8ee1eb2},
journal = {NEURIPS 2024 2024}
}