Nvidia releases a new small, open model Nemotron-Nano-9B-v2 with toggle on/off reasoning

Nvidia releases a new small, open model Nemotron-Nano-9B-v2 with toggle on/off reasoning

In a significant advancement for AI language models, Nvidia has introduced the Nemotron-Nano-9B-v2, a small yet powerful model designed to optimize performance while fitting seamlessly onto a single Nvidia A10 GPU. This release follows a wave of innovative small models, including a recent AI vision model from MIT spinoff Liquid AI and a smartphone-compatible model from Google. The Nemotron-Nano-9B-v2 boasts the ability to toggle AI 'reasoning' on and off, allowing users to choose whether the model self-checks before delivering an answer. With a parameter count reduced from 12 billion to 9 billion, it stands out in its category, demonstrating impressive efficiency and speed—up to six times faster than similar-sized transformer models. Oleksii Kuchiaev, Nvidia's Director of AI Model Post-Training, emphasized that this reduction was tailored specifically for deployment on the widely used A10 GPU. This model supports multiple languages, including English, German, Spanish, French, Italian, and Japanese, as well as several others like Korean and Portuguese, making it versatile for various applications including instruction following and code generation. The Nemotron-Nano-9B-v2 is based on the hybrid Mamba-Transformer architecture, which combines traditional attention layers with selective state-space models to manage lengthy sequences without excessive memory demands. Nvidia's latest offering also features a unique runtime 'thinking budget' management. This allows developers to set limits on the number of tokens spent on internal reasoning, effectively balancing accuracy with latency—a crucial factor for applications such as customer support or autonomous systems. Evaluation results reveal competitive performance against other small-scale models, achieving notable scores on several benchmarks, including a 72.1% accuracy on AIME25 and 90.3% on IFEval. Released under the Nvidia Open Model License Agreement, the model is designed to be enterprise-friendly, allowing developers to use, create, and distribute derivative models without the burden of additional licensing fees or restrictions. This provides a significant advantage for enterprises looking to deploy AI solutions quickly and efficiently. With the Nemotron-Nano-9B-v2, Nvidia is clearly targeting developers who seek a blend of reasoning capabilities and deployment efficiency. The model's features are designed to facilitate experimentation and integration, reinforcing Nvidia's commitment to enhancing AI language model accessibility while maintaining high accuracy and manageable operational costs.

Sources : VentureBeat

Published On : Aug 20, 2025, 03:40

Startups
Meta Exits Renewable Energy Initiative Amid Surge in Natural Gas Projects

In a significant shift, Meta has announced its departure from the RE100 initiative, a prominent corporate renewable ener...

TechCrunch | Jul 23, 2026, 20:05
Meta Exits Renewable Energy Initiative Amid Surge in Natural Gas Projects
AI
Runway Unveils Revolutionary Media Router to Transform Generative Content Creation

Runway is taking a bold step beyond being merely another AI model company; it aims to establish itself as the foundation...

TechCrunch | Jul 23, 2026, 17:45
Runway Unveils Revolutionary Media Router to Transform Generative Content Creation
AI
Cerebras Soars After Strategic Alliance with AMD for AI Innovations

Cerebras Technologies experienced a notable surge in its stock, rising approximately 4% on Thursday following the announ...

CNBC | Jul 23, 2026, 18:55
Cerebras Soars After Strategic Alliance with AMD for AI Innovations
Computing
AMD Partners with Cerebras to Revolutionize AI Inference

In a significant move for the future of artificial intelligence, AMD is shifting its strategy by collaborating with chip...

Business Insider | Jul 23, 2026, 17:50
AMD Partners with Cerebras to Revolutionize AI Inference
AI
Congress Moves to Establish AI Shutdown Protocol Following OpenAI Incident

In a significant legislative response to recent AI vulnerabilities, Representatives Ted Lieu (D-Calif.) and Nathaniel Mo...

CNBC | Jul 23, 2026, 20:05
Congress Moves to Establish AI Shutdown Protocol Following OpenAI Incident
View All News