model roundup

Qwen 3.6

4 items · started 2026-07-17 · closed 2026-07-20

  1. Im developping an agentic app, and using locally qwen 3.6 as the main ai engine. As soon as F5 starts hitting the localhost api of qwen, it gets blocked and rolls back to opus 4.8...

  2. I have three Radeon AI PRO R9700s. I wanted to know if I could tune a quant for this specific hardware instead of just using the generic upstream formats.

  3. so part of what inspired my benchmark post was that it does seem like folks here are generally converging on "unless you are able to operate at very large scales with a lot of system RAM and VRAM, the best model for code work is generally…

  4. DwarfStar Specialized local inference for models that do not fit in memory. A transparent research and co-development fork of antirez/ds4, focused on Metal, adaptive SSD streaming, common 16–64 GB Apple Silicon systems, and measured experi…

← all threads