{
  "video_id": "reddit_1uaebfe",
  "channel_slug": "LocalLLaMA",
  "channel_handle": "r/LocalLLaMA",
  "title": "Best Local Agents - Jun 2026",
  "url": "https://www.reddit.com/r/LocalLLaMA/comments/1uaebfe/best_local_agents_jun_2026/",
  "external_url": null,
  "upload_date": "20260619",
  "published_at": "2026-06-19T21:29:33+00:00",
  "transcript": "A megathread that is overdue! Let's discuss and debate on what the ***best local agents*** available today are\n\n# Prologue\n\nFirst a note on terminology: While most regular users are going to have a general sense of what these are, I think its worth a brief pause to preempt turbulence in the discussion. \n\n- **Agent**: There is no standard/universally agreed upon term that I can find - and rightly so. Its hard to tell if this is a hypecycle buzzword or a new primitive. I think its important to first relate to stuff that already exist and highlight how its new/different. So from that lens, I think it should largely be thought of just another software that takes autonomous/semi-autonomous action based on user input, with the distuinguishing aspect being that it can self determine path/logic and does not require to be pre-programmed (unlike IFTTT, n8n, Apple Shortcuts etc.). This definition largely agrees with /r/AI_Agents's . Or put in another way, we're talking about pi, opencode, hermes etc.\n- **Harness**: I specifically did not use this neologism which seems to be the new buzzword replacing the Agent buzzword, but without any sufficient need. Search/LLMs dont offer a substantative or consensus definition for it either. The best that can eked out is LLM+Harness=Agent. However, I think that's the equivalent of saying Engine+Chassis/Wheels/Steering=Car. So its much more useful to talk about the \"Car\" and thus the titling of this post\n\n\n# The standard spiel: \nstill applies.. \n\nShare what you are running right now and **why**. Given the nature of the beast in evaluating these immature systems (rapidly changing landscape, untrustworthiness of benchmarks, immature tooling, intrinsic stochasticity), please be as detailed as possible in describing your setup, nature of your usage (how much, personal/professional use), how you evaluate etc. Eg: comments like [\"pi is the best\"](https://old.reddit.com/r/LocalLLaMA/comments/1u6njs5/i_think_we_need_a_localharnessllm_or_something/ortxda3/) that doesnt have any substance reduce the quality of the discussion \n\n# Rules\n\n1. Agents must be using open weight models\n2. Agents must be running locally (a.k.a hardware, including VPCs, that you control)\n3. Strongly recommend discussing OSS Agent software but doesn't necessarily have to be so. Why? Claude Code/Codex are relatively the most mature, well understood, largest ecosystem softwares today + they can be used with local models. At least for now we cant ignore the reality that many of us are using those - so its worth allowing at least as a reference point.\n\n\n\n--- Top Comments ---\n\n\n[17 upvotes] pi + llama.cpp + Qwen 3.6 27B Q8 + MTP + ngram with full context on 4x3090s\n\nbecause \"pi is the best\" - it doesn't do bad things with context like OpenCode and it allows me to work with my code without any compromises, this setup is also more responsive than Claude Code (with the cloud) because I don't need to wait every time I type something\n\nI don't really have time to explore other models with this setup because I use existing one for few hours per day (it's addicting)\n\n[8 upvotes] I'm using [clio](https://github.com/SyntheticAutonomicMind/CLIO) \\+ [CachyLLama](https://github.com/fewtarius/CachyLLama) which is my fork that aggressively caches to reduce prompt reprocessing times on low power devices like my AMD APUs.  I'm using this with Qwen 3.6 35B A3B UD Q4 K XL for misc coding work that I don't need to do with a cloud model, system setup work, and other tasks.\n\n[6 upvotes] \"Pi is the best\" is still substantially better than the pedantic arguing over the definitions of words like agent or harness. At least it gives you something to go try.\n \nWhat have the \"erm actually\" *agent means this* or *harness is well defined* people actually contributed here, other than an easy list of blocks to add? (Helps build your own filter to shift through the bs..) \n\n[5 upvotes] Ive thoroughly been enjoying Pi with qwen 3.6 27b and 35b! It is now the first thing I set up when configuring a new VM or PC. And when I feel like I want to do something really weird or I want done 100% on the first try, I use my ChatGPT sub to at least plan it out and qwen 3.6 finish it up ",
  "transcript_chars": 4200,
  "ingested_at": "2026-06-20T01:30:28.263702+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 54,
    "upvote_ratio": 0.88,
    "num_comments": 62,
    "author": "rm-rf-rm",
    "is_self": true
  }
}