NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
video-text retrieval
3 papers
A TRIANGLE Enables Multimodal Alignment Beyond Cosine Similarity
OSKAR: Omnimodal Self-supervised Knowledge Abstraction and Representation
Towards Understanding Camera Motions in Any Video