{
  "video_id": "reddit_1w99zj0",
  "channel_slug": "singularity",
  "channel_handle": "r/singularity",
  "title": "Terence Tao says AI labs’ race to beat math benchmarks is starting to hurt the field. He wants them to compete on new insights instead.",
  "url": "https://www.reddit.com/r/singularity/comments/1w99zj0/terence_tao_says_ai_labs_race_to_beat_math/",
  "external_url": null,
  "upload_date": "20260906",
  "published_at": "2026-09-06T22:15:59+00:00",
  "transcript": "On Aug. 31, Stadlmann posted a preprint lowering the bound on gaps between primes from 246 to 240. Within days, several AI labs were posting their own improvements on social media. Tao responded with an [eight-post thread](https://mathstodon.xyz/@tao/117219548485446992) explaining why he finds this worrying.\n\nHis point is that the number itself was never what mattered most. Going from 246 to 240, or even from 70 million to 246, does little for the rest of mathematics on its own. What mattered was everything developed along the way: Zhang (who was working in a sandwich shop at the time) bringing neglected work on equidistribution back into use, Maynard developing a sieve that became a standard tool, and Polymath8 showing what open collaboration could accomplish. Those advances had applications well beyond the original problem.\n\nTao imagines how things might have played out if today’s AI labs had been around in 2005, when GPY published their near-miss. The labs pour millions into compute, push the bound into the low hundreds, then move on once progress slows. Mathematicians decide the problem has been picked over and look elsewhere. Nobody writes a proper paper or turns the arguments into something people can learn from. An idea like Maynard’s sieve ends up buried in hundreds of pages of AI output that nobody reads. Zhang never gets his moment; Maynard leaves the field.\n\nIn that scenario, the bound improves faster, but mathematics loses out.\n\nIt’s a Goodhart’s law problem: the number becomes the target, and the reasons anyone cared about it get lost. Tao doesn’t think AI has to work this way. He points to the [Erdős problem example](https://terrytao.wordpress.com/2025/12/08/the-story-of-erdos-problem-126/) as a case where collaboration with AI helped advance understanding. His objection is to labs bypassing experts and peer review to rush out a better number. He argues that an approach that was relatively harmless in 2025, when models couldn’t solve whole problems unassisted, has started doing more harm than good in 2026.\n\nHis [proposal](https://mathstodon.xyz/@tao/117221032761877425) is to change what the labs compete over: who can announce a genuinely new mathematical insight first?\n\n\n\n--- Top Comments ---\n\n\n[37 upvotes] This applies to everything and it’s too late now. It’s going to be increasingly ai doing things faster than we can keep up and eventually us out of the loop. He still wants things right for mathematicians but from what I see ai should be focusing on what works for them. They don’t need things explained the same way we do. Surely he should have seen this coming with all the access he’s had, did he think the line wouldn’t keep going up.\n\n[26 upvotes] If this man is worried and some of you think we will have a job... \n\n[12 upvotes] i mean yeah hard to argue with that hes the goat\n\n[9 upvotes] \\>His [proposal](https://mathstodon.xyz/@tao/117221032761877425) is to change what the labs compete over: who can announce a genuinely new mathematical insight first?\n\nThey are already doing this. \n\n[6 upvotes] Idk if there is a good alternative metric though. Or like idk perhaps mathematicians will have to emphasize how the new way or method used to solve it is useful. ",
  "transcript_chars": 3231,
  "ingested_at": "2026-09-07T01:30:16.887289+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 59,
    "upvote_ratio": 0.9,
    "num_comments": 53,
    "author": "Old-School8916",
    "is_self": true
  }
}