{
  "video_id": "reddit_1tnbz23",
  "channel_slug": "LocalLLaMA",
  "channel_handle": "r/LocalLLaMA",
  "title": "Is Qwen3.6 current king for local agentic use?",
  "url": "https://www.reddit.com/r/LocalLLaMA/comments/1tnbz23/is_qwen36_current_king_for_local_agentic_use/",
  "external_url": null,
  "upload_date": "20260525",
  "published_at": "2026-05-25T15:09:33+00:00",
  "transcript": "I've been testing other models but it seems like nothing even come close to Qwen3.6 35B A3B for agentic use. The worse I'd get is a loop sometimes, while Gemma4 produced broken tool calls occasionally and I couldn't even get GLM 4.7 Flash REAP past 2 or 3 messages before it starts looping. All IQ4_NL quants from Unsloth.\n\nI'm wondering if there are better models around the same size (preferably MoE) that I haven't tried yet. I'm using it for Hermes Agent and Pi and it's not perfect, but it's crazy good for a local model\n\n\n\n--- Top Comments ---\n\n\n[137 upvotes] Yes\n\n[37 upvotes] Of course not but for small models, Qwen3.6 27B and 35BA3 are the right choice at the moment.\n\nLocal coding and agentic king is GLM5.1 but most users find that too large to run locally.\n\n[28 upvotes] Yes. I believe it's not even THAT far from DeepSeek v4 Flash.\n\nEDIT: Sorry, I was talking about 27b\n\n[20 upvotes] Qwen is better at coding while I find Gemma better for general user facing. I use both and fine tune both as well! \n\nBig hidden issue is the chat templates cause issues. I redid both the Qwen and Gemma ones for better agentic coding and tool calling fixes. Depending on how you use them there are some weird app side things to take into account with the default chat templates. ",
  "transcript_chars": 1276,
  "ingested_at": "2026-05-26T01:30:01.573322+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 103,
    "upvote_ratio": 0.91,
    "num_comments": 108,
    "author": "HornyGooner4402",
    "is_self": true
  }
}