NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Xiangru Jian
3 papers
University of Waterloo
AlignVLM: Bridging Vision and Language Latent Spaces for Multimodal Document Understanding
Paper2Poster: Towards Multimodal Poster Automation from Scientific Papers
The Underappreciated Power of Vision Models for Graph Structural Understanding