Despite impressive-seeming individual successes, standardized benchmarks consistently show AI having only 1-2% success rates on mathematical problems, and the journal and media attention to successes creates a distorted perception where noise overwhelms signal, making it increasingly important to establish standardized datasets and prevent AI companies from selectively reporting only victories.