Technology2022globalhigh confidence

OpenAI researchers Yuri Burda and Harri Edwards accidentally discovered the 'grokking' phenomenon after arithmetic-training experiments ran for days instead of hours, finding that models which initially memorized example sums eventually learned to add previously unseen numbers.

Notes on verification

Confirmed by original arXiv paper (Power, Burda, Edwards et al.), Wikipedia, and Quanta Magazine reporting; all key details—accidental long training run, arithmetic task, memorization-to-generalization pattern—corroborate consistently. [cascade flags: wikipedia_dropped_other_sources_exist | tier=silver indep_score=0.925 clusters=2 claim_tier=notable]

Sources