E3M: Zero-Shot Spatio-Temporal Video Grounding with Expectation-Maximization Multimodal Modulation

Peijun Bao, Zihao Shao, Wenhan Yang, Boon Poh Ng, A.C. Kot
2/14/2026
DOISemantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. Any repos shown below are weak matches — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{bao2026em,
  title  = {E3M: Zero-Shot Spatio-Temporal Video Grounding with Expectation-Maximization Multimodal Modulation},
  author = {Peijun Bao and Zihao Shao and Wenhan Yang and Boon Poh Ng and A.C. Kot},
  year   = {2026},
  doi    = {10.1007/978-3-031-73010-8_14},
  url    = {https://doi.org/10.1007/978-3-031-73010-8_14},
  journal = {ECCV 2024 2024}
}

Discussion