E3M: Zero-Shot Spatio-Temporal Video Grounding with Expectation-Maximization Multimodal Modulation
Peijun Bao, Zihao Shao, Wenhan Yang, Boon Poh Ng, A.C. Kot
2/14/2026
No confident code match yet
We couldn't find an author-owned or strongly-evidenced community implementation for this paper. Any repos shown below are weak matches — verify before relying on them.
No code implementations found yet.
Know of an implementation? Let us know in the comments below!
@article{bao2026em,
title = {E3M: Zero-Shot Spatio-Temporal Video Grounding with Expectation-Maximization Multimodal Modulation},
author = {Peijun Bao and Zihao Shao and Wenhan Yang and Boon Poh Ng and A.C. Kot},
year = {2026},
doi = {10.1007/978-3-031-73010-8_14},
url = {https://doi.org/10.1007/978-3-031-73010-8_14},
journal = {ECCV 2024 2024}
}