{
  "video_id": "reddit_1u78mx6",
  "channel_slug": "LocalLLaMA",
  "channel_handle": "r/LocalLLaMA",
  "title": "Nex-N2 Pro is the real deal",
  "url": "https://www.reddit.com/r/LocalLLaMA/comments/1u78mx6/nexn2_pro_is_the_real_deal/",
  "external_url": null,
  "upload_date": "20260616",
  "published_at": "2026-06-16T09:29:15+00:00",
  "transcript": "I had dismissed N2 when it was first released due to reports that it performed badly in Openrouter. \n\n\nSo, one good thing came out of the Rio-3.5 model situation: I was so intrigued by Rio's performance that when it came to light that it was just N2 Pro rebranded, it drove me to download and test bartowski's N2 Pro IQ2_S GGUFs. My first N2 tests were breaking due to bugs in the embedded GGUF chat template, but it started working perfectly once I switched to using Rio's chat template.\n\n\nI've been running coding benchmarks on it and super impressed so far. There's a private benchmark where I use it to do some investigation on llama.cpp source code, and it has been passing on it consistently. It is the first model (tested through bartowski's Rio and N2 GGUFs) I can run on my 128G mac that passed on it 100% of the times I tried without hallucinating once, before that only GPT 5.x had this consistency.\n\n\n\n--- Top Comments ---\n\n\n[18 upvotes] Try typing \"hi.\" He'll reply \"hi\" hahaha. I'm using it too, and it works fine in coding/tools calls, but don't use it as a chatbot! Also Rio 1M context.\n\n[9 upvotes] Anyone compared the smaller Nex-N2 mini model with Qwen 3.6 35B-A3B?",
  "transcript_chars": 1184,
  "ingested_at": "2026-06-16T13:30:29.614413+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 54,
    "upvote_ratio": 0.86,
    "num_comments": 37,
    "author": "tarruda",
    "is_self": true
  }
}