Benchmark Contamination Examples?
6.9BRAINROT SCOREBenchmark Contamination When a model's training data accidentally (or not-so-accidentally) includes the actual questions and answers from a benchmark it's later evaluated on, inflating its score without reflecting genuine improved capability.
Origin:Named directly by AI researchers auditing suspiciously high benchmark results, borrowing 'contamination' from the same concept in data science where test data leaks into a training set.
First Seen:2023
Peak Era:2023-2026 (Current Era)
Aura Impact:+15 Aura (Catching and Reporting Real Benchmark Contamination) / -20 Aura (Shipping a Model With Undisclosed Benchmark Contamination)
EXAMPLE USAGE
"Turns out that shockingly high score was just benchmark contamination, the test questions were sitting in the training set."
Want the full breakdown — categories, trend velocity, platform distribution, and community voting on Benchmark Contamination? Visit the full dictionary entry for Benchmark Contamination.