▲ 1 r/kimi

How to prevent Kimi CLI to overexploring and waste tokens?

Kimi K3 used for coding tasks is great! But in the moment i use it as a specific role in my server PM for instance, he starts overdoing/overexploring and burning a lot of tokens unnecessarily.

How have you created a skill or harness that limit the level of exploration Kimi CLI does ? I do use ~yolo mode because most of the projects here are Pet projects with tons of backup, so i don't care if they fuck shit up (preferebly not of course =) )

But i don't have the same issue with Codex or Claude for instance, they remain straight to the point.

Any suggestions? Thanks!

reddit.com
u/3rd_Floor_Again — 20 hours ago

Nature taking over cemetery in Berlin

Alfred died in France during the first world war, his body was never found. Here lies his parents, who passed almost 100 years ago. In the back, you can see several graves, most of the graves of this part of the cemetery don't have a headstone, lost during WW2 or earlier.

u/3rd_Floor_Again — 11 days ago
▲ 109 r/vinyljerk+1 crossposts

Rate my setup

I am a headphone guy. I have a small collection of them, amps, etc. My wife instead of giving me another one for my bday, gave me a turntable and vintage vinyl disks from albums that I like, like a good old Creedence Clearwater Revival 20 greatest hits. I confess I wasn't super hot on vinyl but I have to say I engaging with the albums like my dad (God bless his soul) used to do, actually brought me good childhood memories.

In summary, now I have another rabbit hole to jump after headphones, guess I am buying audio gear since I have no decent speakers. 😅

Any newbie suggestions to start ?

u/3rd_Floor_Again — 11 days ago

Fable 5 (High) >>>> Opus 5 (ULTRA+Workflow)

Anthropic is trying to gaslight us into believing that Opus 5 is almost as good as Fable. That's borderline absurd. Fable just identifies and fixes shit faster. That's it.

reddit.com
u/3rd_Floor_Again — 21 days ago
▲ 1 r/kimi

K3 feels like it's overcomplicating analysis

Decided to test K3 against Fable and GLM 5.2. The one summoning them was GPT 5.6 Sol. I'm using Fable and GLM 5.2 as auditor agents and code advisers that Sol can summon whenever it hits a point in a dev task where it failed 2x.

Playing with K3, I had the impression that, although it felt substantially better than K2.7, it was spending way too many tokens exploring more than it should, even with a tight prompt and a very clear task and boundaries.

I asked Sol to add K3 to the team of auditors/advisers. Sol liked its first work, but decided not to use it anymore due to lack of reliability in resource usage and speed, which was delaying the whole task.

Sol's final verdict on it.

https://preview.redd.it/vjoik0l7j1eh1.png?width=475&format=png&auto=webp&s=6bd28a6ec457315976fcc5135c00011d4dfb2e47

How's the experience been for you guys so far?

reddit.com
u/3rd_Floor_Again — 1 month ago

HD480 Pro, Sennheiser destruiu a concorrência nos closed-backs

Nunca na vida comprei um headphone que não precisava de algum EQ.

A assinatura de som dele é perfeita para mim. Ultra recomendo.

u/3rd_Floor_Again — 2 months ago
▲ 59 r/ZaiGLM

GLM 5.2 consumed quota VERY fast, even on Coding Max Plan

https://preview.redd.it/j21lmfv1778h1.png?width=1135&format=png&auto=webp&s=a3195f0aceb786223f70e316a163bfd8d809e929

Decided to test in a production line the GLM 5.2 on GLM Coding Max Plan vs Opus 4.8 on Claude Max+ (not only on quality but how fast do I run out of Session Tokens).

This was fast. Claude Max+ can keep going for much, much longer. I had to add Claude to keep the tasks going.

At least how the context window is 1M tokens, but with this rate, albeit is a good model, pricing wise on this plan (160 USD) Claude/GPT still delivers more.

reddit.com
u/3rd_Floor_Again — 2 months ago
▲ 14 r/kimi

Will Kimi ever get 1M context window?

One of the reasons why i downgraded Kimi is because for the multi-model production line i built, the model was all the time reaching context limitation and crashin mid-task, even optimizing the instructions to use agents etc. for a specific task, and often it would just ignore the instructions and try to consume the whole files as context, crashing mid way through it.

So was reducing and reducing the usage to the point it became idle. I really enjoy the IDE, although sometimes the model misunderstand what I am asking and start doing stuff he wasn't supposed to do. In general I like it, but reasoning needs to improve a bit and context window size as well.

reddit.com
u/3rd_Floor_Again — 2 months ago

Finally joined the club

I spent a lot of time looking for a closed back that I could do some nice relaxing sessions and I am very happy with my first Sennheiser, the HD 480 Pro.

u/3rd_Floor_Again — 3 months ago

How to use GLM 5.1 with RooCode on VS CODE?

Currently the app only allows me to select GLM 5, it doesn't seem to have a way to manually edit the model to select GLM 5.1.

Am I doing something wrong? Is there a way?

reddit.com
u/3rd_Floor_Again — 3 months ago
▲ 56 r/kimi

I am impressed with Kimi 2.6 > GLM 5.1

It is so f. reliable. Really.

I've been running a team of agents from East to the West side of AI world, similar capabilities, similar tasks, coding.

I feel Z.ai is pretty good, but very expensive, consumes token like there is no tomorrow, even if you are running on glm-4.5-air. The f. concurrency limits is insanely annoying, getting 429 errors all the time, which forces me limit task production, crazy token consumption instead of cacheing correctly, and other problems really made me think that, although they have superb model and really good output quality, it is less competitive when you try to scale it.

On GLM i have the Max Plan and on Kimi Allegro. I feel I am doing way more with Kimi, with a superior model, spending less tokens, than with glm-4.5-air. I even make Kimi 2.6 challenge GLM 5.1 assessments and it always get important gaps. Fuck, it find gaps even from f. Opus 4.7 (which he agrees and very often really praise Kimi work 😃 )

So far, very good value for money, a Sonnet 4.6+++.

reddit.com
u/3rd_Floor_Again — 3 months ago