NVIDIA
NVIDIA's models, often hybrid (Mamba + attention): a much lighter context cache.
Nemotron 3.5 Lightning
August 2026
Nemotron 3
December 2025
Nemotron Nano v2
August 2025