3rd GPU connection with riser
▲ 15 r/LocalAIServers+1 crossposts

3rd GPU connection with riser

Added 3rd GPU and now getting VGA error LED on motherboard. 99% seems riser fault cause it isn't working with any GPU in any slot even when one GPU connected but maybe anybody had something similar with Gigabyte B850 motherboards? PCIe 3.0 x16 30cm riser. Everything was powered on, photo was done before connecting power cable.

Reordered 4.0 20cm.

u/esw123 — 3 days ago

DeepSeek V4 Flash 0731 IQ2_M benchmark for Dual 3060 and 96GB RAM ≈ 3.5 tok/s.

Thanks to the community help I finally launched this llm. LM Studio refused to load weight onto second GPU but Unsloth Studio did so everything was done in there. Not a proper benchmark (used PC in parallel as well) but it gives an idea of ​​the performance from dual 3060 with RAM offloading.

DeepSeek V4 Flash 0731 IQ2_M from Unsloth

CPU: Ryzen 7500F
GPU 0 (PCIe 5.0x16 lane): RTX3060
GPU 1 (PCIe 3.0x1 lane): RTX3060
RAM: 96GB 5600

Prompt: Write me a tetris game.

Result (copy from the summary):
Prompt eval: 2.96s
Prompt speed: 3.0 tok/s
Generation: 960.02s
Speed: 4.5 tok/s (PowerShell shows 3.5 tok/s, don't know why it shows 4.5 tok/s, single GPU with RAM offload output was around 3 tok/s so I trust PowerShell metrics more)
Tokens: 4,338
First token: 2.96s
Cache hits: 1
Total: 963.32s
Chunks: 4318

Wattmeter is on the way but my estimate is around 130W total system power draw (with 2 monitors connected but they are not taken into account) Each GPU used 30-40W with 0.9V undervolt. Task took around 16 minutes to complete, consumed around 35 watts and cost 0.0059 euros.

Update:
OS Windows 11.

Actually math is showing 4338/960.02=4.52 tok/s, don't know why PowerShell showed 3.5 most of the time. I ran one more request with web search and it gave 4.7 tok/s. Updated 3.5 -> 4.5 tok/s.

u/esw123 — 20 days ago

Installation and setup

I want to try unsloth studio but can not initiate installation already, how to fix this:

Using this tutorial: https://unsloth.ai/docs/new/studio/install#windows

PS C:\WINDOWS\system32> irm https://unsloth.ai/install.ps1 | iex

  🦥 Unsloth Studio Installer (Windows)
  ────────────────────────────────────────────────────

  winget         available
  python         Python 3.13 already installed
                 preserving existing environment for rollback...
                 previous environment preserved for rollback
  venv           creating Python 3.13 virtual environment
                 C:\Users\___\.unsloth\studio\unsloth_studio
  gpu            NVIDIA GPU detected
                 installing PyTorch (https://download.pytorch.org/whl/cu130)...
                 installing unsloth (this may take a few minutes)...
  unsloth        2026.7.6 installed
  setup          running unsloth studio setup...
Refusing to run Unsloth inside System32 as it will lead to Errors.
cd to a normal working directory and try again.
[ERROR] unsloth studio setup failed (exit code 1)
                 restoring previous environment after failed install...
                 restored previous environment
unsloth studio setup failed (exit code 1)
At line:153 char:9
+         throw $Message
+         ~~~~~~~~~~~~~~
    + CategoryInfo          : OperationStopped: (unsloth studio ...d (exit code 1):String) [], RuntimeException
    + FullyQualifiedErrorId : unsloth studio setup failed (exit code 1)
u/esw123 — 20 days ago

Best way to organize 6 GPUs

Need to organize 6 GPUs in closed case, can somebody share their approach, please? From what I found the best way is to buy rig frame but all of them are open. 4U case nowhere to fit and only big case that I found is new Phanteks Enthoo Elite Server but it is around 400 euro. If I go for open air rig, how to protect and cover it? Maybe you can share any tips what to look for, additional space for second PSU later or built-in HDD cage? Thanks.

reddit.com
u/esw123 — 27 days ago

Dual 3060 upgrade path

I've started building my rig and it turned out 24GB VRAM is not enough for my tasks (inference only without offload to RAM). What will be the best way for an upgrade for dual 3060 - add used 3090 for 700-800 euro or add 2 more barely used 5060Ti 16GB for the same price. Slower newer but 56GB vs 48GB total. Don't know how nvfp4 is usable actually.

reddit.com
u/esw123 — 1 month ago

Dual 3060 MoE loading issue.

I had one RTX3060 and was able to use all models like Qwen3.6 27B, Qwen3.6 35B and Qwen3.5 122B with CPU offload. I've added second 3060 and now only dense model can be loaded as usual, 35B loaded only once and 122B refuses to load at all. What setting should I change or where to look for potential problem solving solutions?

Update: Seems 12 layers in VRAM is the limit for dual 3060 with qwen3.5-122b-a10b. Reduced to 12 and model loaded. 35b-a3b loaded as well after PC restart and when I changed GPU priority back to 3060 installed in PCIe 5.0x16 from the second GPU installed in PCIe 3.0 x1 (deviceID: 1 and then deviceID: 0).

u/esw123 — 1 month ago

Second drive

Do I really need to upgrade my main drive with OS or I just can save a little bit by keeping 480GB SSD for OS and Metashape and just add SN7100 2TB as second drive for projects? Will it somehow affect processing times or only Windows and Metashape starting time?

reddit.com
u/esw123 — 1 month ago

Pigtailed 3090 safe to buy?

Is it safe to buy used 3090 which was pigtailed (2x8-pin)? Don't know how long.

reddit.com
u/esw123 — 2 months ago

How to improve RAM offload?

I have only 12GB VRAM (RTX3060) but have enough RAM to run Qwen3.6 27B Q4 with offload. Something tells me that it won't achieve maximum performance but why DRAM speed is only around 30GB/s (HWiNFO data) during inference with dual channel 5200 RAM? TG is 3.12 tok/sec with 18K tokens result.

I expected slow speed, but can't understand where is the bottleneck, is it how LM Studio works or I need better CPU (I have 7500F). Of course dual 3090 will do the work, but it is what is for now.

Tried smaller prompt with 6 CPU threads, Q8 KV cache, 37 GPU offload, got TG 4.95 tok/sec and bandwidth was 30-35GB/s.

u/esw123 — 2 months ago

How much you paid for AI Max+ 395 128GB in Europe?

I am looking at one right now and can't understand why mini pc is around €4000 while Asus ProArt PX13 is available for €3000. Both with 128GB memory while laptop is on the go platform with extra battery and display. Is it because of TDP limits or is it a good deal for €3000?

reddit.com
u/esw123 — 2 months ago

How good AI395+ is for Agisoft?

Has anyone tried Ryzen AI 395 for Agisoft Metashape? I found a good deal with 128GB RAM but not sure about it. How much worse it would be vs regular 96-128GB RAM PC with RTX5080 and R9 9950X. Or it will be fine, or there will be need for NVIDIA GPU via Thunderbolt to complete work faster? Aim is to complete projects with 2000-4000 images, higher is better if possible.

reddit.com
u/esw123 — 3 months ago