I am using 20x plan. I am seeing a lot of hallucination and scope drift, and extremely slow execution.
model
GLM-5.2
huggingface.co/zai-org/GLM-5.2 ↗
1267198 downloads4634 likestext-generationtransformers
from the model card
GLM-5.2 👋 Join our WeChat or Discord community. 📖 Check out the GLM-5.2 blog and GLM-5 Technical report. 📍 Use GLM-5.2 API services on Z.ai API Platform. 🔜 Try GLM-5.2 here. [Paper] [GitHub] Introduction We're introducing GLM-5.2, our latest flagship model for long-horizon tasks. It marks a substantial leap in long-horizon task capability over its predecessor GLM-5.1 and, for the first time, delivers that capability on a solid 1M-token context. GLM-5.2's new capabilities include: Solid 1M Context: A solid 1M-token context that stably sustains long-horizon work Advanced Coding with Flexible Effort: Stronger coding capabilities with multiple thinking effort levels to balance performance and latency Improved Architecture: We propose IndexShare, which reuses the same indexer across every four sparse attention layers, reducing per-token FLOPs by 2.9× at a 1M context length. We also improve GLM-5.2’s MTP layer for speculative decoding, increasing the acceptance length by up to 20% Pure Open: An MIT open-source license — no regional limits, technical access without borders Benchmark |Benchmark|GLM-5.2|GLM-5.1|Qwen3.7-Max|MiniMax M3|DeepSeek-V4-Pro|Claude Opus 4.8|GPT-5.5|Gemini 3.1 Pro| |:---|:---:|:---:|:---:|:---:|:---:|:---:|:---:|:---:| |Reasoning||||||||||| |HLE|40.5|31|41.4|37|37.7|49.8|41.4|45| |HLE (w/ Tools)|54.7|52.3|53.5|-|48.2|57.9|52.2|51.4*| |CritPt|20.9|4.6|13.4|3.7|12.…
discussions
recent items
Hallucinations and severe scope drifts inspite of planning. (www.reddit.com via reddit) DGX Spark, cluster of 4 (www.reddit.com via reddit) Does anyone have a first-hand experience with four Sparks cluster, and how much of an upgrade is it comparing to just two considering the available models? While there's plenty of noise for the smaller models (Qwen) and our older king Deep…
Getting GLM-5.2 NVFP4 Post-Training off the ground (patronus.ai via hn) Getting GLM-5.2 NVFP4 Post-Training off the ground The goal was deceptively simple to state: take GLM-5.2, a 744B-parameter mixture-of-experts model quantized to 4-bit NVFP4, attach a bf16 LoRA adapter, and train it with reinforcement lear…
What We Learned Moving Our Agent Loops from Anthropic to GLM (getunblocked.com via hn) What We Learned Moving Our Agent Loops from Anthropic to GLM Why we moved most of Unblocked's agent traffic from Claude Opus to GLM 5.2, what the blind A/Bs and the ledger actually showed, and what broke on the way. TL;DR: We moved most of…
Engy – Verified LLM Inference (engy.ai via hn) Verified inference means cryptographic proof that the exact open model you requested produced your output, not a cheaper or quantized stand-in. Run frontier open models like GLM-5.2, billed by the token.
Kimi K3 and GLM 5.2 can create undetectable malware for $2 (www.incalmo.ai via hn) The danger frontier: low-cost, evasive, abundant malware As part of Incalmo’s mission to make AI safely ubiquitous, we do safety research on the frontier cyber capabilities of models. Recently, to help anti-virus systems stay ahead of the…
Coinbase Switches to Chinese AI Models GLM and Kimi, Cuts AI Spending by 50% (mlq.ai via hn) Coinbase Switches to Chinese AI Models GLM and Kimi, Cuts AI Spending by 50% - Coinbase has defaulted engineers to GLM 5.2 from Zhipu and Kimi 2.7 from Moonshot AI through its internal LLM gateway, cutting AI spending by nearly 50% [1] - G…
We built the new fastest API for GLM-5.2 (www.baseten.co via hn) A month ago, GLM-5.2 was released. As part of our day-zero support, we built the fastest API in the world for GLM-5.2, with peak speeds of 280 tokens per second and average speeds around 100 tokens per second.
Sakana: Fugu Ultra vs. GLM 5.2 (runtimewire.com via hn) This wasn’t a photo finish. Sakana: Fugu Ultra controlled the matchup on practical writing and coding tasks, while GLM 5.2 showed flashes of polish in a couple of narrower instruction-following spots.