NVIDIA Nemotron-Labs TwoTower: 2.42x Faster Inference, No Retraining Required
NVIDIA Nemotron-Labs TwoTower retrofits a pretrained 30B AR model into a diffusion decoder running 2.42x faster at 98.7% quality. Open weights, commercial license, ...
Tools, open source, and developer productivity