started · updated
NVIDIA launches Nemotron 3.5 Lightning to boost AI agent efficiency
NVIDIA has released Nemotron 3.5 Lightning, a 30-billion-parameter Mixture-of-Experts (MoE) model optimized for high-efficiency AI agent workloads. By utilizing proprietary NVFP4 4-bit floating-point precision and Quantization-Aware Distillation (QAD), the model achieves up to four times the throughput of full-precision versions while significantly reducing memory requirements from 66GB to 22GB. To support this ecosystem, NVIDIA also introduced NeMo Switchyard, an open-source routing library designed to direct AI requests to the most suitable models based on latency, cost, and quality.
In the broader industrial landscape, companies are integrating these advancements into physical and manufacturing environments. Advantech is deploying Edge AI and AI Agents through its IWS platform to enable autonomous sensing and analysis on factory floors. Similarly, Hocheon Technology is focusing on integrating humanoid robots and automated systems to address labor shortages, emphasizing practical industrial applications over purely simulated models.
NVIDIA is also expanding its strategic influence through significant capital investments. The company has pledged up to $105 billion in support for OpenAI’s data center projects in Ohio and is working with financial institutions to treat GPUs as a new asset class, facilitating third-party investment in AI infrastructure.
Entities
Advantech · Hechuan Technology · Hocheon Technology · NeMo Switchyard · Nemotron 3.5 Lightning · Nvidia · OpenAI