{
  "video_id": "reddit_1u8kr2o",
  "channel_slug": "LocalLLaMA",
  "channel_handle": "r/LocalLLaMA",
  "title": "We need a 80-160B model urgently.  The unified memory device market needs more Models.",
  "url": "https://www.reddit.com/r/LocalLLaMA/comments/1u8kr2o/we_need_a_80160b_model_urgently_the_unified/",
  "external_url": null,
  "upload_date": "20260617",
  "published_at": "2026-06-17T19:55:07+00:00",
  "transcript": "Hello guys,   \nI will keep myself short.   \n**There are so many people that have a lot but not enough of \"slow\" RAM.**   \nAnybody with a Apple Device with >96GB  \nAnybody with a Ryzen AI 395 Device with >96GB  \nAnybody with a DGX Spark  \nEven people with RTX 6000 Pros or 4x3090s or other configurations.  \nOr People with 128GB DDR4/5 RAM\n\n**Yet the models that came out in the last 3 months**   \nwere particulary made for high speed low capacity machines   \n(27B Qwen, 31B Gemma)  \nor the other extreme, massive models   \n(GLM 5.2, Deepseek V4 Pro, Kimi 2.7, Mimo 2.5 Pro, MiniMax M3) \n\n**We people with unified memory devices or other 80-128GB configurations**  \nhave to either use older models that are not great at all currently as the frontier has expanded.  \n(Glm 4.5 Air, GPT OSS 120B, Qwen 3.5 122B, Nemotron 3 Super 120B, Qwen 3 Coder Next 80B)\n\nOr we have to use small models due to our slow bandwidth RAM/VRAM  \n(Qwen 3.6 35B or Gemma 4 26B)\n\n**We need something in the range of 100B 10B Sparse.** Something that people with a AMD 9700 AI Pro or a Rtx 3090/5090 and 64GB Vram could use. Something that DGX Spark Users, Ai395+, Apple Users, etc.\n\nSomething like Gpt OSS 120B V2, Gemma 4 122B, Qwen 3.6/3.7 122B, GLM 5.2 Air, Deepseek V4 Mini with 100B, Mimo 2.5 Mini with 100B or anything similar to that class of models. Or heck even a Qwen 3.6 Coder 80B would be something people would love. \n\nI really hope we are gonna get something - else I am left with Qwen 3.5 122B on my Spark for now. \n\nCheers.\n\n\n\n--- Top Comments ---\n\n\n[1 upvotes] Your post is getting popular and we just featured it on our Discord! [Come check it out!](https://discord.gg/PgFhZ8cnWW)\n\nYou've also been given a special flair for your contribution. We appreciate your post!\n\n*I am a bot and this action was performed automatically.*\n\n[177 upvotes] I too would like this. Who's manager do I need to speak to? lol\n\n[68 upvotes] TL;DR: we have GPUs sitting between 64-128GB doing nothing useful because every model is either too small to bother with or too big to fit. Someone please cook a 100B sparse MoE.",
  "transcript_chars": 2091,
  "ingested_at": "2026-06-18T01:30:24.620962+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 277,
    "upvote_ratio": 0.88,
    "num_comments": 174,
    "author": "Storge2",
    "is_self": true
  }
}