▲ 14 r/ollama

How to optimise local AI for lots of RAM but not a lot of VRAM?

Im running a Ryzen 7 5700x, a 3080ti (12GB) with 64GB of RAM. I’m still new to Local AI, and I’ve tried it in the past but none of the previous generations of AI have been good enough for my specific niche use case. Yesterday I tried qwen 3.8 27b and it looked really promising. However on my 3080ti it offloaded to RAM slightly and turned the model agonisingly slow (unsure of exact decode or output speed).
I didn’t mess with any of the config and was running Ollama. Is there anything I can do to take advantage of my RAM and increase speeds?

reddit.com
u/Top_Drink8324 — 1 day ago

Anyone try running qwen 3.8 27b on a 3080ti at any quant?

I was thinking about trying it. I understand that it won't fully fit but im curious about quality with speed.

reddit.com
u/Top_Drink8324 — 5 days ago

Can someone figure out what i'm doing wrong?

I use claude for programming, however the free tier has very real and obvious limitations when it comes to how long i can use it for. I then realized my pc was actually sufficient for a local LLM and decided to try it. However every single one i've tried just feels incredibly dumb to me. I'm using qwen3-coder30b-A3B and it always,always get extremely simple issues just wrong. Ill ask it to change the colour of something to red for example and it will instead retrieve random services it doesn't even need, change the colour to red, and proceed to also change the material of it. It constantly messes up scope too.
Is it just a config/prompting error? Or is this just a natural limitation of local AI?

reddit.com
u/Top_Drink8324 — 1 month ago
▲ 3 r/ollama

Wondering if someone could help me realize what i'm doing wrong

I use claude for programming, however the free tier has very real and obvious limitations when it comes to how long i can use it for. I then realized my pc was actually sufficient for a local LLM and decided to try it. However every single one i've tried just feels incredibly dumb to me. I'm using qwen3-coder30b-A3B and it always,always get extremely simple issues just wrong. Ill ask it to change the colour of something to red for example and it will instead retrieve random services it doesn't even need, change the colour to red, and proceed to also change the material of it. It constantly messes up scope too.
Is it just a config/prompting error? Or is this just a natural limitation of local AI?

reddit.com
u/Top_Drink8324 — 1 month ago