Nvidia’s Nemotron 3.5 Lightning and NeMo Switchyard aim to reduce AI agent costs by routing tasks across specialized and frontier models.
Nvidia's Nemotron 3.5 Lightning pairs with its NeMo Switchyard router, which reassigns models mid-task and cuts task costs to a third in Nvidia's own tests.
Nvidia (NVDA) is expanding its Nemotron 3 open model family with Nemotron 3.5 Lightning, which it calls the highest-efficiency model in its class for long-running agentic AI workloads. Nemotron 3.5 ...
Soaring AI infrastructure costs and model pricing, combined with uncertain returns on investment, threaten to stall enterprise adoption. To make enterprise AI spend a bit more manageable, Nvidia this ...
2don MSN
Nvidia releases new open model
Company says Nemotron 3.5 Lightning can help enterprises save on AI token costs ...
Nvidia releases Nemotron 3.5 Lightning and NeMo Switchyard to give enterprise AI capability options - SiliconANGLE ...
Nvidia's new open model to process agent workloads faster. The open-source library NeMo Switchyard includes building blocks for model routers.
NVIDIA Nemotron Lightning delivers 4x throughput for AI agents. This 30B parameter model handles repetitive execution tasks with high speed and low cost.
Nemotron 3.5 Lightning delivers up to four times faster output and 30% faster agentic task completion than other models in its class.
Some results have been hidden because they may be inaccessible to you
Show inaccessible results