Are video generators the secret to better computer vision?

Researchers at Google Deepmind found a new way to use video generators. They taught a model to understand depth and object shapes in images. The system performed just as well as existing tools but used much less data.
Most of this training happened on synthetic videos instead of real photos. This proves that video generators already understand how the physical world works. They do not just copy pixels but learn the logic behind them.
This discovery changes how we look at AI development. It suggests that if we want smarter machines, we should teach them to predict video. We might have had the missing pieces for better vision all along.
Comments (0)
No comments yet. Be the first!
More AI news
NewsNew AI Tool Fixes Mistakes in Financial Research
Researchers developed a new framework to stop AI from learning from its own bad data in financial trading models.
NewsPerplexity adds local AI to Mac apps
The new update lets your Mac handle some AI tasks directly on your computer.
NewsMeta Releases Muse Image Model on Fal
Meta has launched a new AI model on the Fal platform that plans and edits images on its own.