NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Han Cai
3 papers
NVIDIA
Jet-Nemotron: Efficient Language Model with Post Neural Architecture Search
Sparse VideoGen2: Accelerate Video Generation with Sparse Attention via Semantic-Aware Permutation
Win Fast or Lose Slow: Balancing Speed and Accuracy in Latency-Sensitive Decisions of LLMs