On August 12, 2026, NVIDIA announced the release of its new open-source model, Nemotron 3.5 Lightning. This event marks a significant step in the democratization of artificial intelligence, aimed at reducing costs and simplifying the deployment of AI agents. The new model, designed for rapid task execution, promises to fundamentally change the approach to automating routine operations in the corporate sector and among independent developers.
Mixture-of-Experts Architecture: More Power at Lower Costs
The key feature of Nemotron 3.5 Lightning is its unique architecture. Although the model has 30 billion parameters, thanks to the Mixture-of-Experts (MoE) technology, it activates only 3 billion parameters to process each token. This engineering solution allows for performance comparable to powerful systems, while utilizing computational resources characteristic of significantly more compact models.
Such optimization makes the model an ideal candidate for performing repetitive operations: verifying results, executing commands, and formatting data. While larger neural networks, such as Nemotron 3 Ultra, continue to handle strategic planning and solving complex tasks, Lightning takes on the "dirty work," freeing up powerful compute nodes for more intelligent functions.
Automation Tools and Local Deployment
To effectively manage task flows, NVIDIA released the NeMo Switchyard library. This tool analyzes the complexity of incoming requests and automatically redirects simple instructions to be processed by the Lightning model, saving resources and time. Developers emphasize that this significantly reduces computational costs.
Special attention has been paid to technology accessibility. The model is optimized for operation in popular environments like OpenClaw and Hermes Agent and supports the NVIDIA NemoClaw security stack. Thanks to tools like NeMo Automodel and NeMo Megatron Bridge, deployment is faster and cheaper. Furthermore, the model can be run locally on DGX Spark and Jetson platforms, as well as on personal computers with GeForce RTX 5090 graphics cards, making it accessible to a wide range of users.
Open Source and Developer Ecosystem
NVIDIA continues its course on openness by providing code, weights, and datasets under the OpenMDW-1.1 license. The model comes with the open Nemotron-RL Agentic Terminal Pivot dataset, specifically prepared for training AI coders. This solution is designed to stimulate community development and accelerate the creation of new applications based on AI agents.
In the context of the current date, August 12, 2026, this release appears as a response to the growing demand for efficient and accessible automation tools. NVIDIA's move not only strengthens its market position but also sets new standards for the industry, where efficiency and accessibility become key success factors.