NVIDIA Unveils Diffusion-Based Nemotron-TwoTower-30B Model
Reported by the original publisher: NVIDIA has released Nemotron-TwoTower-30B-A3B-Base-BF16, an unusual diffusion-based language model built from the Nemotron 3 Nano 30B-A3B backbone.. Analysis and context written by TickrWire.
NVIDIA released Nemotron-TwoTower-30B-A3B-Base-BF16, a diffusion-based language model using a two-tower architecture with parallel token generation, built on the Nemotron 3 Nano 30B-A3B backbone.

- Nemotron-TwoTower-30B-A3B-Base-BF16 is a diffusion-based language model, not a traditional autoregressive one.
- It uses a two-tower architecture: a frozen autoregressive context tower and a diffusion denoiser tower for parallel token generation.
- The model retains 98.7% of the performance of its backbone (Nemotron 3 Nano 30B-A3B).
- Built with BF16 precision, it targets efficiency and scalability improvements.
- Released as part of NVIDIA's Nemotron series, though details are sparse beyond the Reddit announcement.
NVIDIA has unveiled Nemotron-TwoTower-30B-A3B-Base-BF16, a novel diffusion-based language model that diverges from traditional autoregressive generation. Instead of generating tokens sequentially, it employs a frozen autoregressive context tower paired with a diffusion denoiser tower that fills blocks of tokens in parallel. This approach aims to improve efficiency and scalability while retaining 98.7% of the original model's performance. The model is built on the Nemotron 3 Nano 30B-A3B backbone and uses BF16 precision for inference.
Introduces a novel parallel token generation approach for language models, potentially improving inference speed and scalability for large models.
Could enable more efficient deployment of large language models, reducing compute costs and latency for inference tasks.
Signals NVIDIA's continued innovation in model architectures, which may influence market positioning in AI infrastructure.
Demonstrates an alternative to traditional autoregressive models, relevant for research in diffusion-based language generation.
Highlights NVIDIA's push beyond standard LLM architectures, though practical impact remains to be seen.
- diffusion-based language model
- A model that generates text by iteratively refining noisy token sequences, rather than predicting tokens sequentially.
- two-tower architecture
- A model design with separate specialized components (e.g., context and denoiser towers) working in tandem.
- BF16
- A 16-bit floating-point format used for efficient model inference and training.
- autoregressive generation
- A model that generates tokens one at a time, conditioned on previously generated tokens.
AI bias estimate: Neutral reporting of a technical release; limited context beyond Reddit source. (Automated estimate, not a definitive judgement.)
LLMQwen3.8-27B: A Deep Dive Into Qwen's Newest Vision-Language Powerhouse
LLMAlibaba's Qwen team releases Qwen 3.8 models with open weights under the Apache 2.0 license
Meta’s ‘open’ AI, and a $250M deal gone very wrong
LLMZ.ai Ships GLM-5.3 Without Retraining the Base Model: Better at Complex Coding and Long-Horizon Tasks
LLMGemini 3.7 Flash lands with coding gains and undercuts its three-week-old predecessor's price by 50%
AI ToolsMeta AI’s new Mac app wants you to talk to your apps
Meta released a new Mac application that lets users control apps and dictate text using voice commands powered by its Muse Spark AI model.
New White House strategy clarifies military tech priorities: undersea, outer space and AI - Breaking Defense
The White House released a new strategy prioritizing military investments in artificial intelligence, space systems and undersea technologies to counter emerging threats.
AI in an iron grip: How dictatorships use artificial intelligence to strengthen their rule - theins.press
A new report examines how authoritarian governments deploy AI for surveillance, censorship, and propaganda to reinforce their power.
Stripe, OpenRouter finally strike a deal - Banking Dive
Stripe and OpenRouter have partnered to integrate Stripe's payment processing with OpenRouter's AI model aggregation platform.
How one Philadelphia school is using AI to strengthen student learning, not replace teachers - CBS News
A Philadelphia school is integrating AI tools to support teachers and improve student outcomes, focusing on collaboration rather than replacement.
Exclusive-How a Texas student blew the whistle on a rogue AI hacking attempt - The Mighty 790 KFGO
A Texas student uncovered an AI-powered hacking attempt targeting local systems, prompting a swift law enforcement response.