{
  "video_id": "reddit_1tt7t5t",
  "channel_slug": "singularity",
  "channel_handle": "r/singularity",
  "title": "The new benchmarks like DeepSWE now show a very big gap in proprietary models and open source",
  "url": "https://www.reddit.com/r/singularity/comments/1tt7t5t/the_new_benchmarks_like_deepswe_now_show_a_very/",
  "external_url": null,
  "upload_date": "20260531",
  "published_at": "2026-05-31T21:19:08+00:00",
  "transcript": "Before we could only see a few points between closed and open source models. Hopefully open source can catch up a bit more. At the moment it is quite disappointing.  \n\n\nhttps://preview.redd.it/prwafwsghj4h1.png?width=1448&format=png&auto=webp&s=04b2656474065e6bd3c15c244d585c542f8f526d\n\n\n\n--- Top Comments ---\n\n\n[50 upvotes] Basically, it's something all of us heavy users already knew. Unfortunately, open source models are about 6–8 months behind. But bots, people with incentives, and subreddits with weird cults will tell you that’s not the case because they don’t do anything professional and just mess around with code or simple stuff\n\n[9 upvotes] The real question imo is are these gaps a big enough pain point for the average non-enterprise user? \n\nConsumers increasingly want more bang for their buck. Are these improvements worth 3x-50x the costs? I can use a Chinese model that is at \"good enough\" level far more on the same budget then I can with GPT/Opus.\n\n[7 upvotes] I don't understand how Gemini 3.5 flash scores so high. I really can't get that quality out of it.",
  "transcript_chars": 1080,
  "ingested_at": "2026-06-01T01:30:27.666879+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 58,
    "upvote_ratio": 0.92,
    "num_comments": 21,
    "author": "sitytitan",
    "is_self": true
  }
}