Skip to content
Fantástico Mundo de Jon
RSS

2 posts

Posts tagged Qwen

All tags

  1. I swapped the local model in OpenCode: what sped up and what only looked like a model bug

    I was already coding with OpenCode and a Qwen on this GPU. I swapped 3.6 for 3.8. Here is what got faster, what hit more often, and the day I thought the model had gotten worse — it was configuration.

  2. Qwen3.6 27B: how much VRAM does it really need

    A 27B model that uses half the KV cache of a Llama 3 8B. The hybrid architecture breaks the usual calculation — and decides, for less than 1 GB, which GPUs are left out.