# Nvidia Releases a 31.6-Billion-Parameter Open Model That Activates 3.6 Billion Per Token

Nemotron 3.5 Lightning uses a hybrid Mamba-Transformer design, runs on a single H100, and scores 24 on the Artificial Analysis Intelligence Index against 15 for its predecessor.

- Published: 2026-08-12T06:11:01.127Z
- Canonical: https://polylog.news/ai/2026-08-12/nvidia-releases-a-31-6-billion-parameter-open-model-that-act
- Publisher: Polylog (AI desk)
- Section: tech
- Sources: [Ollama](https://ollama.com/blog/nemotron-3-5-lightning), [NVIDIA Blog](https://blogs.nvidia.com/blog/local-ai-open-source-models-agents-nemotron/), [Hugging Face](https://huggingface.co/nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16), [The Decoder](https://the-decoder.com/nvidias-open-weight-nemotron-3-5-lightning-prioritizes-speed-over-maximum-intelligence/)

Nvidia released Nemotron 3.5 Lightning, an open-weight mixture-of-experts model with 31.6 billion total parameters and roughly 3.6 billion active per token, published in both BF16 and NVFP4 four-bit checkpoints and available through Ollama…

This story is for subscribers. Read it in full at https://polylog.news/ai/2026-08-12/nvidia-releases-a-31-6-billion-parameter-open-model-that-act (subscription information: https://polylog.news/pricing).