TGViewer
Data eXplore : Data Science, ML, Big Data, LLMs and AI Security Data eXplore : Data Science, ML, Big Data, LLMs and AI Security @dataxplore · 582 subscribers
Post #1914 205
NVIDIA has introduced a new open family of Nemotron 3 models

1️⃣ Nemotron 3 Nano
Universal model for reasoning and chat, focused on local deployment.

Key characteristics:
- MoE architecture: 30B parameters total, ~3.5B active
- Context up to 1 million tokens
- Hybrid architecture:
- 23 Mamba-2 + MoE layers
- 6 attention layers
- Balance between speed and quality of reasoning

Requirements:
- About 24 GB of video memory is needed for local deployment

The model is well suited for long dialogues, document analysis, and reasoning tasks

An interesting example of how MoE and Mamba are actually starting to reduce hardware requirements while maintaining context scale and quality.


2️⃣ Nemotron 3 Super & 3️⃣ Nemotron 3 Ultra
significantly surpass Nano in scale - by about 4 times and 16 times respectively. But the key point here is not just the size of the models, but how NVIDIA managed to increase power without a proportional increase in inference cost.

NVFP4 and the new Latent Mixture of Experts architecture are used for training Super and Ultra. It allows for four times more experts to be used at the same inference cost. Essentially, the model becomes "smarter" due to a more flexible selection of experts, rather than constantly activating all parameters.

Additionally, Multi-Token Prediction is used, which accelerates training and improves the quality of reasoning on long sequences. This is particularly important for agentic and multi-agent scenarios, where models work with long context and complex decision chains.


A good signal for the industry. Release, Guide, GGUF, lmstudio

#AI #LLM #NVIDIA #Nemotron3 #OpenSource #MachineLearning

••••••••••••••••••••••••••••••••••••••
🤖 Data Science, ML & Big Data with @DataXplore
More from @dataxplore
  1. Oct 1, 2026Monitoring and debugging "silent" token drift in LLM pipelines: frequency analysis of embe…
  2. Sep 30, 2026Online detection of feature collisions in TDA transformation When using Topological Data A…
  3. Sep 18, 2026Am going to announce something big (for me, it's really big) on October 11, 2026.
  4. Sep 14, 2026Post #2188
  5. Aug 31, 2026I joined a Russian community on Telegram. They share some Russian startup and technology u…
  6. Aug 22, 2026Post #2185
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →