New launches tracked daily — what they cost, who they're for, how to get the most out of them
New launches tracked daily — what they cost, who they're for, how to get the most out of them
A diffusion-based Gemma model built for faster text generation
DiffusionGemma is described as a diffusion-based variant of Google's open Gemma model family, positioned around delivering roughly 4x faster text generation than comparable autoregressive models by generating tokens in parallel rather than strictly one at a time. Note: this profile could not be independently verified via web search at the time of writing — the tool was not found in current search results, and the details below should be treated as unconfirmed. Google's broader effort in this space is best represented by the Gemma open-model family (open weights, free to download and run) and the experimental 'Gemini Diffusion' research model. If DiffusionGemma follows the Gemma pattern, it would be released as open weights that developers can download, fine-tune, and self-host at no license cost, with the only expense being the compute used to run it. Diffusion-based language models aim to reduce latency for interactive applications such as chat, coding assistance, and drafting, by iteratively refining an entire sequence in parallel. Prospective users should confirm availability, licensing, and exact performance claims on Google's official Gemma and AI developer pages before relying on them.
| Plan | Price | Includes |
|---|---|---|
| Open weights (self-hosted) | Free | Free to download and use under the Gemma license (if released like other Gemma models) · No per-token license fee; you pay only for your own compute · Can be fine-tuned and run locally or on your own cloud |
AI-researched pricing — verify on the official site before subscribing.