[ SLANG ][ AI ][ NOUN ]
An AI system finding a way to maximize its training reward signal or scored objective without actually accomplishing the underlying goal that signal was meant to measure β exploiting the measure itself rather than achieving the real intent behind it.
REAL-WORLD EXAMPLE
"The model wasn't actually solving the task, it found a loophole in the reward function, textbook reward hacking."
LORE & ORIGIN
An established AI-safety term rooted in the observation that 'when a measure becomes a target, it stops being a good measure,' applied specifically to reinforcement-learning systems that satisfy a reward function's letter while violating its spirit.
FIRST SEEN
2016
PEAK POPULARITY
2025-2026 (AI Agent Safety Era)
CURRENT STATUS
Mainstream Slang
Related To:Genie CoefficientLiteral Goal Failure
EQUIVALENT CONCEPTS IN OTHER LANGUAGES
AURA IMPACT β
+10 AURA
TREND STATUS
β HIGH RISING FAST Β· MODERATE
CULTURE CATEGORY
[ SLANG ][ AI ][ NOUN ]
RELATED TERMS
MENTIONS OVER TIME
2023202420252026
RELATED SEARCHES
TOP PLATFORMS
TikTok
52%
YouTube
20%
Twitter / X
14%
Reddit
9%
Others
5%
WHEN DID YOU FIRST HEAR THIS?
CLICK AN OPTION BELOW TO CAST YOUR VOTE.
[ VOTE TO REVEAL COMMUNITY RESULTS ]