German AI team corrects benchmark error in new model

The team behind the Soofi S model recently published an updated report. They discovered that some questions from a major science test were part of their training set by mistake. Community members noticed the issue after looking through the public data.
Because the model had already seen these questions, its performance results were inaccurate. This happens when a computer accidentally practices the exact test it is supposed to take later.
In response, the developers removed that benchmark from their official results. They recalculated everything to provide a fair picture of how well the model actually works. They are now working to ensure this does not happen again.
Comments (0)
No comments yet. Be the first!
More AI news
NewsNew AI Tool Fixes Mistakes in Financial Research
Researchers developed a new framework to stop AI from learning from its own bad data in financial trading models.
NewsPerplexity adds local AI to Mac apps
The new update lets your Mac handle some AI tasks directly on your computer.
NewsMeta Releases Muse Image Model on Fal
Meta has launched a new AI model on the Fal platform that plans and edits images on its own.