{
  "video_id": "reddit_1wcsj6v",
  "channel_slug": "LocalLLaMA",
  "channel_handle": "r/LocalLLaMA",
  "title": "New tensor type layouts for my GGUF uploads",
  "url": "https://www.reddit.com/r/LocalLLaMA/comments/1wcsj6v/new_tensor_type_layouts_for_my_gguf_uploads/",
  "external_url": null,
  "upload_date": "20260910",
  "published_at": "2026-09-10T19:05:03+00:00",
  "transcript": "Hey all, long time no post.\n\nFigured I'd pop my head in to point you towards a blog post I just published about research I had performed and changes I'm making to the shape of models I post, you can read it here:\n\nhttps://huggingface.co/blog/bartowski/per-tensor-layout-maps-for-gguf-quantization\n\nI won't try to claim \"Pareto frontier\" or \"best models in the world\", but I will say from tests the new shapes look to be better across the board than what I was posting before, so I'm really happy with where it came out, and I hope to not be done yet either :)\n\nhttps://cdn-uploads.huggingface.co/production/uploads/6435718aaaef013d1aec3b8b/Ufz9TXQlKFxVHdocVoZIw.png\n\nIf anyone has any questions let me know!\n\n\n\n--- Top Comments ---\n\n\n[8 upvotes] >If a file that starts with Q3\\_K\\_ is mostly non-Q3\\_K tensor types and sits above 5 bits per weight, something has gone wrong..\n\nThis\n\n[7 upvotes] This looks very cool! However, whenever I try to tell people on reddit that quants smaller than Q4 can actually be useful, they don't believe me.\n\n[4 upvotes] Welcome back! You said \"better across the board\" and attached a no-Pareto-frontier disclaimer, and honestly that's refreshing around here.\n\n[3 upvotes] Cool, when will it be merged into llama-cpp?",
  "transcript_chars": 1250,
  "ingested_at": "2026-09-11T01:30:07.421461+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 55,
    "upvote_ratio": 1.0,
    "num_comments": 17,
    "author": "noneabove1182",
    "is_self": true
  }
}