u/DanManREAL_GRIND

Rtx 5080 +5060 :)

Rtx 5080 +5060 :)

My precious 💖

Currently Running 3 Local Models (llama.cpp, ALL 3 run in Claude Code):

SEAT 1 - LONG-CONTEXT DAILY DRIVER. Current incumbent: Qwen3.8-27B-Heretic-Q4_K_M. Measured at 64K: 30.80 tok/s on an empty context and 22.98 tok/s after the same approximately 61K-token prompt. Its job is 96K–112K terminal-log and evidence work.

SEAT 2 — RARE QUALITY MODE. Current incumbent: Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-Q5_K_M.gguf, exact size 17.91 GiB. It already loaded at context 65536, generated 1500 tokens at 19.09 tok/s, processed a 61,387-token prompt, then generated another 122 tokens at 17.07 tok/s with truncation=0.

SEAT 3 — MTP SPEED MODE. Current incumbent: AEON-ULTIMATE-UNCENSORED-IQ4_XS.gguf with an embedded MTP head.

u/DanManREAL_GRIND — 3 days ago