{
  "video_id": "reddit_1u6cwvn",
  "channel_slug": "ClaudeAI",
  "channel_handle": "r/ClaudeAI",
  "title": "Stop burning your Opus 4.8 Max limits. Use this prompt to slash token costs.",
  "url": "https://www.reddit.com/r/ClaudeAI/comments/1u6cwvn/stop_burning_your_opus_48_max_limits_use_this/",
  "external_url": null,
  "upload_date": "20260615",
  "published_at": "2026-06-15T10:39:24+00:00",
  "transcript": "If y'all are running ~~Fable 5~~ Claude Opus 4.8 on the Max plan, you already know it’s an absolute mofo beast for engineering workflows. But let’s be real, because it defaults to high-effort adaptive thinking, it also chews through token caps and session limits like crazy. It loves to over-explain things, add polite conversational filler, and rewrite a massive file just to change two lines of code.\n\nTo stop wasting generation limits, I’ve been using a highly compressed prompt that forces it to skip the fluff and maximize high-density output. Just copy-paste this into your custom instructions, project knowledge, or the start of your chat:\n\n`Launch subagents. Output only the modified or requested code block. Do not provide line-by-line explanations, setup guides, introductory, concluding remarks, or markdown commentary unless explicitly asked. Adopt an ultra-concise, high-density communication style.`\n\nTelling it to **\"Launch subagents\"** immediately triggers the model’s dynamic workflow architecture. Instead of the primary model burning massive reasoning tokens to plan out a sprawling task, it offloads task execution efficiently to parallel sub-processes. The biggest token saver by far is commanding it to **\"Output only the modified or requested code block.\"** Opus Max has a terrible habit of rewriting an essay-script just to show a minor modification. This constraint completely shuts that down, forcing it to give you only the specific diff or function you actually asked for, which slashes your output token costs to almost zero (not really zero).\n\nOn top of that, explicitly stripping explanations, setup guides, and polite concluding remarks eliminates all the redundant filler you already know anyway. Finally, demanding an **\"ultra-concise, high-density communication style\"** forces Claude’s adaptive thinking mechanism to heavily COMPRESS its syntax, ensuring every single token returned carries maximum technical signal.\n\n\n\n--- Top Comments ---\n\n\n[55 upvotes] Do not use a hail mary prompt. It is a science, do evaluate what kind of work you are doing and which model is suitable for that. You can instead mention about your persona in your [claude.md](http://claude.md) and the output would be more aligned with that. Its not a code monkey\n\n[9 upvotes] The \"output only the code block\" part is the real MVP here, Opus rewriting a whole file for a 2-line change is a genuine plague. Skeptical that \"launch subagents\" does anything magic though.\n\n[7 upvotes] I fucking hate Opus 4.8 so fucking much. ",
  "transcript_chars": 2531,
  "ingested_at": "2026-06-15T13:30:12.697020+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 56,
    "upvote_ratio": 0.66,
    "num_comments": 30,
    "author": "BananaIsles",
    "is_self": true
  }
}