model roundup

Opus 4.8

4 items · started 2026-08-26 · closed 2026-08-30

  1. I have a project I am working on, a scientific research project involving around 40 journal articles, that are complex, utilizing many concepts, and inter-related ideas, which I am using Claude to analyze and write about. I have been using…

  2. Megathread for discussing the release of GLM-5.3-Flash. Quants Fine-Tunes & Abliterations Chat Templates Inference Server Support & Configuration Experiences, Benchmarks & Model Comparisons We'll try to clean up future duplicates around th…

  3. Opus 4.8 and Opus 5 tied on strict functional score across 25 matched Stet tasks, but their paired artifacts differed: Opus 4.8 had the lower footprint on 20 tasks, while Opus 5 ran more shell commands on 18 and more tests on 15.

  4. I don't know but after opus 4.8 the models seem to use a lot of jargon and verbose language. I hope this is not the case with me only.

← all threads