{
  "video_id": "reddit_1tz5ffp",
  "channel_slug": "LocalLLaMA",
  "channel_handle": "r/LocalLLaMA",
  "title": "You don't need a GPU to run gemma-4-26B-A4B",
  "url": "https://www.reddit.com/r/LocalLLaMA/comments/1tz5ffp/you_dont_need_a_gpu_to_run_gemma426ba4b/",
  "external_url": null,
  "upload_date": "20260607",
  "published_at": "2026-06-07T07:24:27+00:00",
  "transcript": "I've been running LLMs on my old potato i5-8500 with 32GB of RAM and \\*no GPU\\* for awhile now, running up to 12B dense models which run slow but perfectly useable. But this Gemma-4-26B-A4B simply flies on this CPU - only machine using Koboldcpp on Linux.\n\nThat's right, an old used $150 desktop computer is running state of the art LLMs with something like 7 T/s. Yeah, go ahead and scoff. You can brag about your super-rig that costs more than a used car, but I'm bragging about a crappy old desktop I bought of ebay running the same thing that costs less than a night out.\n\nI keep thinking about buying a GPU but it's beginning to look like it might not be necessary. These smaller models are amazing without a GPU.\n\n\n\n--- Top Comments ---\n\n\n[53 upvotes] I would not go to such extremities, but you have a point. You CAN run reasonably decent models on consumer-grade rigs. The Big Tech is trying way too hard to convince regular people the only way to use AI is cloud, otherwise you literally need a server rack in your bedroom.\n\n[40 upvotes] Yeah, the Gemma 26BA4 only have 4B of active parameters, so it's a lot easier to run even on a potato, provided you could fit the model in RAM.\n\n[23 upvotes] That is true in the same way that you don’t need a car to go to a place 1000 miles away when you can walk after all. ",
  "transcript_chars": 1322,
  "ingested_at": "2026-06-07T13:30:05.921662+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 121,
    "upvote_ratio": 0.84,
    "num_comments": 94,
    "author": "JackStrawWitchita",
    "is_self": true
  }
}