u/GuaranteePurple4468

Context Shift causing significant slowdown?

Not sure if this is just my system or what, but I find that if I enable Context Shift it significantly increases the VRAM usage of the model I am using, almost guaranteeing it overflows into memory. The same happens with Smart Context.

EG, using a 12.8gb Gemma 4 K_S quant with 48 layers set, 32k context (Q5 kv cache), with FF, SWA and Smart Cache gets my total vram usage up to about 14.3gb including windows processes.

However, changing that to use Context Shift instead of SWA, and suddenly my entire 16gb VRAM is fulled and an extra 11gb is getting loaded into memory, completely tanking the t/s to unuseable levels.

Is there any way around it at all? The loss of performance is just too big for me to justify using it currently.

reddit.com
u/GuaranteePurple4468 — 3 days ago

Kimi 2.5 going away end of August

Well damn, there goes my favorite model on Openrouter. Anyone got any alternatives?

My experiences:

Kimi 2.6 seems to be pure thought incarnate and never actually gets to the reply part, and Kimi K3 is just far out of my budget.

GLM 5.2 is too much of a "yes man" and is extremely predictable, prompts can lower the agree-ability but you can still notice the way it tries to avoid conflict.

Mimo 2.5 does seem to feel fresh compared to the GLM series, but it makes some stupid choices sometimes (to be fair, it does also make some pretty fun choices at times). The Pro version just feels too stiff.

Gemma 31B is decent when it comes to storytelling, but could use a bit more work on it's logic. Pretty similar to Mimo 2.5 in my eyes.

Deepseek V4 and Pro both just feel too passive for me. I struggle to get decent length responses from them and they are hesitant to move the story along or introduce plot hooks, really needs you to go out of your way to point it in a direction. Would be fine if you are the one driving the plot though. Also the new costs a too high.

reddit.com
u/GuaranteePurple4468 — 4 days ago

OW1 - Co-op lag even though both players on same local network

Been having an issue with Co-Op and hoping someone can assist.

We have been experiencing bad delays (eg: noticeable latency of multiple seconds when opening inventories, seeing each other in different locations, enemy hits landing at different times etc) when playing co-op.

This makes no sense to me as

  1. We are both on the same local network, with gigabit ethernet connections
  2. According to the Wiki Outward is P2P so there shouldn't be online servers affecting the connection

Does anyone know how to remove this delay, or maybe force the P2P connection to go through LAN instead of the internet connection (since that seems to be what it is doing)?

reddit.com
u/GuaranteePurple4468 — 3 months ago

Any way to increase gamma beyond the base?

Could use some help here.
Game is so dark I cannot even see the steepness of a hill directly in front of me, and I am already set to max brightness.

Might be due to my screen having an extremely low minimum darkness/ true black level or whatever it is called so I don't know how the below would look to other people.

On my screen the hills, pillars, walls etc in the below screenshots are pitch black (as in, zero texture visible). A bandit can be swinging a mace in my face and I would not even know he was there without locking on.

Currently stuck in a bandit camp and cannot find the exit or a hill I can climb over due to how dark things are (I don't have a torch or lantern, or any way to craft them).

Any config files I can edit, or mods I can install to increase the brightness to useable levels for my monitor?

I can only tell this is a hill because of the outline at the top, but it's so dark it looks like a flat black wall.

Only directly next to the camprifre is visible. The pillars and whatever is next to them is just solid black, I cannot see the stairs or anything on the path leading to the campfire, even seeing my own character is difficult.

reddit.com
u/GuaranteePurple4468 — 3 months ago

Kimi K2.6 (free) on Openrouter - huge amount of failures to respond

Anyone else having this issue with Kimi K2.6 Free on Openrouter?

At least 90% of all my messages get an instant failure "Empty response from API" and I have to retry so many times just to proceed with any of my chats.

Finding it really weird, because the last time this happened was with the older Deepseek R1 model during the time it was losing providers, and it was clearly visible on Openrouter as downtime. In this instance though Openrouter has no indication of any issues with K2.6 free so really confused what is happening.

Can this be caused by prompts at all, is there anything I can do from my side to make it more stable?

Or is this 100% just on the provider side?

reddit.com
u/GuaranteePurple4468 — 3 months ago

Gemma 4 31B issues with reasoning and completion API question

So I'm struggling to wrap my head around this and hoping someone smarter than me can help.

I am trying to get reasoning enabled in Gemma 4 31B (eg: Mero-Artemis-31B-v0.3.1 ), but can't seem to figure out **how** to do that.

I am loading the model using KoboldCPP, with the latest SillyTavern update, and I have it set in my Chat Completion template to Request Reasoning, and have Reasoning Effort set to "High", but the model simply does not output any reasoning at all (nothing hidden in the KoboldCPP terminal either, it just is not thinking at all).

Do I have to use Text Completion instead?
And if so, where do I actually add my prompt/preset details? I cannot see anywhere in Text Completion to add my preferred wall of text.

u/GuaranteePurple4468 — 3 months ago

I've seen a few people mention you can set your hardware on Huggingface and it can tell you what models you can run, but for the life of me I cannot find where to do that.

Could someone be kind enough to point me in the right direction?

reddit.com
u/GuaranteePurple4468 — 4 months ago

Anyone have experience with Koboldcpp and troubleshooting it? I can't find any logs so no idea why this is happening.

I have a 16gb Amd Radeon RX 6800 with 80gb desktop memory.

The steps I have done:

  1. Downloaded koboldcpp-1.104 from the YellowRose Rocm github.
  2. Downloaded a model from huggingface (SuperGemma4-31b-abliterated.Q4_K_M.gguf) 17.4gb.
  3. Opened the Koboldcpp exe file and left it on default settings, selected the model and clicked launch.

The result is... nothing.
The exe just closes, then... nothing.
No errors, nothing I can see in the background in task manager, just... nothing.

Tried messing around with context sizes etc but seems like they all do the same.

reddit.com
u/GuaranteePurple4468 — 4 months ago