Google DeepMind researchers developed DiffusionGemma, a text diffusion model that leverages an existing model, Gemma 4, to significantly reduce training costs. By parallelizing token generation, DiffusionGemma achieves a higher output rate than its predecessor, reaching 1,500 tokens per second. However, the model's performance still lags behind the original autoregressive model, particularly in tasks requiring reasoning. The development demonstrates a more efficient approach to building text diffusion models, which could have implications for large-scale natural language processing applications.
Google Improves Text Diffusion Model Efficiency
Original source
Read the full story at The Decoder →This is an original summary written by Rouagent News. The reporting belongs to The Decoder. Follow the link for their full article.
