Report Ads

Nvidia Releases New Open Model Nemotron 3.5 Lightning and Switchyard Router to Cut Enterprise Costs

Nvidia
From gaming to AI, Nvidia drives visual computing innovation. [TechGolly]

Key Points:

  • Nvidia expanded its open model ecosystem by launching the new Nemotron 3.5 Lightning model and NeMo Switchyard router.
  • The lightweight Nemotron 3.5 Lightning features 30 billion parameters utilizing a mixture-of-experts architecture designed for fast agentic workflows.
  • NeMo Switchyard functions as an open-source model routing library that automatically selects appropriate models to minimize enterprise token expenses.
  • The hardware giant aims to help companies maintain corporate intellectual property security and manage computational costs across complex workloads.

The artificial intelligence landscape is witnessing a strategic push toward efficient, customizable software architectures. Nvidia announced the release of a new open model named Nemotron 3.5 Lightning alongside an open-source model routing tool called NeMo Switchyard. This dual release is designed to give enterprise organizations greater control over artificial intelligence deployment, data privacy, and operational token costs across multi-agent systems.

The newly unveiled Nemotron 3.5 Lightning is a 30-billion-parameter mixture-of-experts model optimized specifically for high-volume, long-running agentic tasks. Unlike massive frontier models that consume significant computing power for routine assignments, this efficient model activates only a fraction of its parameters at any given time. According to company specifications, the model delivers up to four times faster output speed and achieves 30% faster task completion compared to alternative open models in its weight class.

To maximize the efficiency of these mixed model environments, Nvidia introduced NeMo Switchyard. This open-source routing library acts as an automated traffic controller for enterprise intelligence workflows. Rather than routing every consumer prompt or back-end task through expensive, high-capacity frontier models, the router evaluates complexity in real time. It dynamically assigns simpler sorting and retrieval tasks to compact models like Lightning while reserving heavier reasoning workloads for larger systems, balancing speed, latency, and financial cost.

Security and data privacy served as primary drivers behind the development of these tools. Many corporate clients operating within specialized domains—such as cybersecurity, healthcare, and advanced material science—hesitate to share proprietary data with external cloud providers due to fears of intellectual property leakage or market competition. By utilizing open-weight models, enterprises can train and fine-tune software locally using proprietary data sets, ensuring that sensitive corporate information stays protected.

Industry analysts note that this software rollout reinforces the company’s hardware dominance by making artificial intelligence deployment more affordable and scalable for mainstream businesses. As enterprise customers look for ways to optimize their multi-million-dollar technology budgets, offering high-efficiency open models and intelligent routers provides a practical pathway toward sustainable, long-term automation.

Newsroom
Newsroom
Al Mahmud Al Mamun leads the TechGolly Newsroom team. He served as Editor-in-Chief of a world-leading professional research Magazine. Rasel Hossain is supporting as Managing Editor. Our team is intercorporate with technologists, researchers, and technology writers. We have substantial expertise in Information Technology (IT), Artificial Intelligence (AI), and Embedded Technology.