Video, Ergo Genero: Unifying Video Tasks via Spatiotemporal Analogy
Adapting video models to new tasks typically requires dedicated data curation and fine-tuning. While visual analogy provides a training-free alternative by specifying tasks in-context, it remains restricted to the image domain. To explore whether analogy-based methods can unify diverse video tasks and generalize to out-of-distribution scenarios, we introduce ViGeo, a framework that extends visual…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.