August 18, 2026Running DeepSeek V4 Flash Q4_K_XL at ~100 tok/s prompt processing on 4× RTX 3060 12GBAugust 18, 2026·reddit.com
August 18, 2026If the weights never change, is it really recursive self-improvement?August 18, 2026·reddit.com
August 18, 2026OpenCode overrides the samplers for Qwen models to the wrong valuesAugust 18, 2026·reddit.com
August 18, 2026Any ~4B models with decent programming knowledge? NO VIBECODE/ONE SHOTAugust 18, 2026·reddit.com
August 18, 2026What would make an uploader-run refusal table independently reproducible?August 18, 2026·reddit.com
August 18, 2026DeepSeek V4 Flash 0731 on Strix Halo: draft model, n_max sweep, and a launch line that actually helpsAugust 18, 2026·reddit.com
August 18, 2026Don't ignore llama.cpp RPC with old hardware. Results of a 5070 Ti and 1080 Ti over gigabit ethernet: it's actually functional.August 18, 2026·reddit.com
August 18, 2026Qwen 3.8 27B xhigh vs medium small comparison (+ others for fun)August 18, 2026·reddit.com