Google turns Gemma into a faster text generator

Google DeepMind took their Gemma model and turned it into a diffusion model. This method uses only a small fraction of the original training budget. It is a much cheaper way to build new AI tools.
This new version creates 256 tokens at once instead of writing one word at a time. This change allows it to hit speeds of 1,500 tokens per second. It is a big step forward for generation speed.
The model is not perfect yet. It still performs worse than the original version on complex logic and reasoning tasks. It shows that speed and quality are still hard to balance.
Comments (0)
No comments yet. Be the first!
More AI news
NewsGoogle AI Changes Its Search Advice After Bias Complaints
Google updated its search tool after it incorrectly told users to call emergency services based on a person's nationality.
NewsWhy AI Is Still Failing at Simple Tasks
Researchers gave an AI five thousand dollars to grow, but it could not even open a bank account.
NewsEnovis to Buy eCential Robotics
Enovis is expanding its surgical tech business by purchasing French company eCential Robotics.