cd /news/artificial-intelligence/nvidia-gives-away-ai-models-for-free… · home topics artificial-intelligence article
[ARTICLE · art-109355] src=cryptobriefing.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

Nvidia gives away AI models for free to boost GPU sales

Nvidia released Nemotron 3.5 Lightning, a 30-billion-parameter open-weight AI model, on August 11, 2026, as part of a strategy to give away AI models for free to boost GPU sales. The chipmaker also provides free hosted inference for over 100 models via build.nvidia.com APIs, and earlier in 2026 launched the 550-billion-parameter Nemotron 3 Ultra and Dynamo 1.0 software that enhances GPU performance by up to 7x.

read2 min views1 publishedAug 24, 2026
Nvidia gives away AI models for free to boost GPU sales
Image: Cryptobriefing (auto-discovered)

Via nvidia.com

The chipmaker's open-weight Nemotron family and free inference APIs are designed to get developers hooked on its hardware ecosystem

Nvidia is handing out AI models like free samples at Costco. The logic is the same, too: once you try the product, you’ll come back to buy the hardware that runs it best.

The company’s latest release, Nemotron 3.5 Lightning, is a 30-billion-parameter model that launched on August 11, 2026. It uses a Mixture of Experts (MoE) architecture with roughly 3 billion active parameters, meaning it can run on a single GPU while supporting context windows up to 1 million tokens. That’s a serious amount of capability for a model you can download from Hugging Face without paying a dime.

The razor-and-blades playbook, supercharged #

Nvidia now provides free hosted inference for over 100 AI models through its build.nvidia.com APIs. Developers can access these models without swiping a credit card, test them against their own workloads, and build applications on top of them.

Earlier in 2026, Nvidia dropped the Nemotron 3 Ultra, a 550-billion-parameter behemoth, alongside Dynamo 1.0, software that reportedly enhances GPU performance by up to 7x.

Why open models matter for the GPU business #

The AI industry has largely split into two camps. On one side, you have companies like OpenAI and Anthropic building closed, proprietary models behind API paywalls. On the other, you have Meta with its Llama series and now Nvidia pushing open-weight alternatives that anyone can download, modify, and deploy.

The models are released under permissive licenses, covering entire families like Nemotron and Cosmos. They’re optimized for Nvidia-specific hardware formats like NVFP4, a quantization format designed for Nvidia chips.

In August 2026, Nvidia also introduced NeMo Switchyard, a tool that intelligently routes tasks across different models to reduce inference costs.

Disclosure: This article was edited by Editorial Team. For more information on how we create and review content, see our

Editorial Policy.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @nvidia 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/nvidia-gives-away-ai…] indexed:0 read:2min 2026-08-24 ·