Recommended settings for Qwen3.6 35B oQ4e + MTP
I need recommendations on the model settings to get optimal performance in oMLX. I am using the latest 0.6.2 version. I have a Macbook Pro with M1 Max and 64GB ram.
- What context window should I use? Should I use 128k or 256k?
- Which one will deliver better performance, TurboQuant or Lightning MTP?
- If I use Lightning MTP, what would be the recommended context window?
I am using omlx and local models particularly for agentic coding via VS Code and OpenCode.