Embodied Intelligence Observer

Developing Nemotron 3.5 Lightning NVFP4 with QAD Using NVIDIA Model Optimizer

Industry

Source: NVIDIA 技术博客Publish time unverified

Teams customize their models to hit their targets for latency, speed, memory, and compute. With the open NVIDIA Nemotron family of models, developers can find... Teams customize their models to hit their targets for latency, speed, memory, and compute. With the open NVIDIA Nemotron family of models, developers can find the right-sized model for their needs. The new Nemotron 3.5 Lightning NVFP4 checkpoint, for example, preserves accuracy while unlocking up to 4x faster throughput.

Developing Nemotron 3.5 Lightning NVFP4 with QAD Using NVIDIA Model Optimizer | Embodied Intelligence Observer