AI Future Leakage: The Silent Flaw Breaking How We Test Whether Machines Can Predict the Future
Northwestern University researchers showed that the standard method for testing AI models for 'future leakage' falsely accuses models: four out of five flagship models failed a test they could not fail. All questions were dated after the models' training cutoff, so answers could not have been memorized.
- 4 of 5 flagship AI models falsely failed a data leakage test
- All test questions were dated after models' training cutoff
- Scientists mathematically proved the standard test method is flawed
Read next
AI