Close Menu

    Subscribe to Updates

    Get the latest creative news from infofortech

    What's Hot

    How to Claim Your Cut of Apple’s $250 Million Siri Settlement

    September 22, 2026

    Poitras Center to fuel early careers of 50 young scientists dedicated to psychiatric disorders research | MIT News

    September 22, 2026

    Who owns the customer relationship when an agent does the buying? – GeekWire

    September 22, 2026
    Facebook X (Twitter) Instagram
    InfoForTech
    • Home
    • Latest in Tech
    • Artificial Intelligence
    • Cybersecurity
    • Innovation
    Facebook X (Twitter) Instagram
    InfoForTech
    Home»Artificial Intelligence»NVIDIA launches open model family for agentic AI
    Artificial Intelligence

    NVIDIA launches open model family for agentic AI

    InfoForTechBy InfoForTechJanuary 21, 2026No Comments3 Mins Read
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr WhatsApp Email
    NVIDIA launches open model family for agentic AI
    Share
    Facebook Twitter LinkedIn Pinterest Telegram Email


    The Nemotron 3 lineup – comprising Nano, Super, and Ultra – delivers leading performance for multi-agent AI systems, combining advanced reasoning, conversational, and collaborative capabilities. The models leverage a hybrid Mamba-Transformer mixture-of-experts (MoE) architecture, providing best-in-class inference throughput while supporting context lengths of up to 1 million tokens.

    Nemotron 3 Nano, the smallest model, is optimized for cost-efficient inference and tasks such as software debugging, content summarization, AI assistant workflows, and information retrieval. Despite possessing 30 billion total parameters, it intelligently activates only about 3 billion per token. With a unique hybrid MoE design, Nano achieves up to 4× higher token throughput than its predecessor and reduces reasoning-token generation by 60%, all while maintaining superior accuracy. Early benchmarks show Nano outperforming comparable open models like GPT-OSS-20B and Qwen3-30B on reasoning and long-context tasks.

    Nemotron 3 Super and Ultra extend these capabilities for high-volume collaborative agents and complex AI applications, incorporating innovations such as latent MoE, a hardware-aware expert design that increases model quality without sacrificing efficiency, and multi-token prediction (MTP), which enhances long-form text generation and multi-step reasoning. Both larger models are trained using NVIDIA’s NVFP4 format, enabling faster training and reduced memory requirements.

    All Nemotron 3 models are post-trained using multi-environment reinforcement learning (RL), enabling them to handle tasks spanning mathematical and scientific reasoning, competitive coding, instruction following, software engineering, chat, and multi-agent tool use. The models also support granular reasoning budget control at inference time, allowing developers to fine-tune computational resources while maintaining accuracy.

    NVIDIA has also released a comprehensive suite of datasets, training libraries, and evaluation tools, including over three trillion tokens of pretraining and reinforcement learning data, the NeMo Gym and NeMo RL open-source libraries, and the Nemotron Agentic Safety Dataset for real-world safety evaluation.

    The Nemotron 3 family is designed to empower developers, startups, and enterprises to build specialized AI agents transparently and efficiently. Nano is available today through Hugging Face, NVIDIA NIM microservices, and major cloud and AI platforms including AWS, Google Cloud, and Microsoft Foundry. Super and Ultra are expected to launch in the first half of 2026.

    Early adopters such as Accenture, ServiceNow, Perplexity, and Palantir are already integrating Nemotron 3 models into AI workflows for manufacturing, cybersecurity, software development, media, and enterprise operations.

    With Nemotron 3, NVIDIA is working on a new standard for efficient, accurate, and open AI models. This will allow developers to scale agentic AI applications from prototype to enterprise deployment while maintaining transparency, cost-efficiency, and state-of-the-art performance.

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    InfoForTech
    • Website

    Related Posts

    Poitras Center to fuel early careers of 50 young scientists dedicated to psychiatric disorders research | MIT News

    September 22, 2026

    Why AI Adaptation, Not Adoption, Is the Real Work Ahead

    September 22, 2026

    AI in Business: Overcoming Deployment Challenges

    September 21, 2026

    A new chapter for MIT Reads | MIT News

    September 19, 2026

    AI that knows its limits

    September 18, 2026

    How OpenAI’s GPT-6 Astra Can Help You Build Presentations

    September 17, 2026
    Leave A Reply Cancel Reply

    Advertisement
    Top Posts

    A Billionaire-Backed Startup Wants to Grow ‘Organ Sacks’ to Replace Animal Testing

    March 23, 2026375 Views

    DoJ Disrupts 3 Million-Device IoT Botnets Behind Record 31.4 Tbps Global DDoS Attacks

    March 20, 202641 Views

    Mayiduo spent S$1M to produce his movie. It broke even & that’s a win in S’pore.

    March 31, 202634 Views

    How is Luckin Coffee expanding rapidly in S’pore while keeping its coffee so cheap?

    April 23, 202622 Views
    Stay In Touch
    • Facebook
    • Twitter
    • Pinterest
    • Instagram
    • YouTube
    • Vimeo
    Advertisement
    About Us
    About Us

    Our mission is to deliver clear, reliable, and up-to-date information about the technologies shaping the modern world. We focus on breaking down complex topics into easy-to-understand insights for professionals, enthusiasts, and everyday readers alike.

    We're accepting new partnerships right now.

    Facebook X (Twitter) YouTube
    Most Popular

    A Billionaire-Backed Startup Wants to Grow ‘Organ Sacks’ to Replace Animal Testing

    March 23, 2026375 Views

    DoJ Disrupts 3 Million-Device IoT Botnets Behind Record 31.4 Tbps Global DDoS Attacks

    March 20, 202641 Views

    Mayiduo spent S$1M to produce his movie. It broke even & that’s a win in S’pore.

    March 31, 202634 Views
    Categories
    • Artificial Intelligence
    • Cybersecurity
    • Innovation
    • Latest in Tech
    © 2026 All Rights Reserved InfoForTech.
    • Home
    • About Us
    • Contact Us
    • Privacy Policy

    Type above and press Enter to search. Press Esc to cancel.

    Ad Blocker Enabled!
    Ad Blocker Enabled!
    Our website is made possible by displaying online advertisements to our visitors. Please support us by disabling your Ad Blocker.