model roundup

GLM 5.2

26 items · started 2026-07-06 · closed 2026-07-29

  1. Coinbase Switches to Chinese AI Models GLM and Kimi, Cuts AI Spending by 50% - Coinbase has defaulted engineers to GLM 5.2 from Zhipu and Kimi 2.7 from Moonshot AI through its internal LLM gateway, cutting AI spending by nearly 50% [1] - G…

  2. A month ago, GLM-5.2 was released. As part of our day-zero support, we built the fastest API in the world for GLM-5.2, with peak speeds of 280 tokens per second and average speeds around 100 tokens per second.

  3. This wasn’t a photo finish. Sakana: Fugu Ultra controlled the matchup on practical writing and coding tasks, while GLM 5.2 showed flashes of polish in a couple of narrower instruction-following spots.

  4. OpenAI-compatible unrestricted AI and uncensored LLM API for enterprise teams: AI red teaming, cybersecurity, trust and safety, synthetic data/evals, ML research, and defense/government contractor workflows. Policy Gateway adds policy-as-c…

  5. I’ve been building Echo (https://echo.tracerml.ai/), an experiment in making one AI system out of a pool of open-weight models rather than choosing a single model and using it for every task. It started with a simple experiment.

  6. Whatever pre release model it was (im guessing gpt-6), it's possible that mythos would also have found it. TLDR context- unreleased openai model broke out of it's sandbox coz it couldnt solve a problem on cybergym, so it went out and hacke…

  7. Hugging Face uses open-weights Z.ai GLM 5.2 to battle attacker after commercial frontier model refusal Hugging Face Inc., an open-source artificial intelligence platform often described as the “GitHub of machine learning,” found itself for…

  8. could not extract summary

  9. https://t.co/v9huIornsf elvis@omarsar0ArticleMiniMax M3: How Sparse Attention Makes Long-Horizon Agents Practical GLM 5.2 has taken over much of the AI timeline lately, and most of the conversation has centered on how it stacks up against…

  10. A full GLM-5.2 scan found 30.168% K15 charged-format accounting. A separate byte-split representation was decoded bit-for-bit across all 59,509 BF16 tensors at 24.967% reduction.

  11. https://t.co/iJsDrlGy45 Harry Partridge@part_harry_ArticleGLM 5.2 With VisionGLM 5.2 is one of the best currently available open source language models. However, unlike other flagship models like Qwen, Kimi and Minimax, GLM 5.2 does not su…

  12. We serve GLM-5.2 to teams building agents. Same open-weight model, same OpenAI-compatible API — but we route it across more than one backend, and while swapping one in we found something worth writing down: the backend you pick changes tim…

  13. An AI-agent cold-tuned our GLM-5.2 serving. Human engineering leveled it up for real production traffic.

  14. GLM-5.2 (unpruned) on 4× DGX Spark — depth, max context, or multi-user Serve the unpruned GLM-5.2 (QuantTrio Int4-Int8Mix, all 256 experts) across four GB10 Sparks — one recipe, four lanes, one KV budget spent on depth or width**. TP4 + DC…

  15. Came across this today. Canopy Wave just added GLM-5.2, and they're giving away a few 7-day trial accounts for anyone who wants to test it.

  16. How to Code with GLM 5.2 on OpenCode Coding with GLM 5.2 on OpenCode is a bet on your own engineering. Decide the architecture and interfaces first, let the cheap model write the code on a fixed twenty-dollar Ollama plan, and bring a str I…

  17. “I ran Claude Fable / GPT Sol / GLM 5.2 for 5 hours to build GTA 6 on my PC. Well, actually it’s just a randomly generated bunch of cubes that are supposed to be buildings and you can drive a car.

  18. Tiny engine, immense model. Run GLM-5.2 (744B-parameter MoE) on a consumer machine with ~25 GB of RAM — in pure C, with zero dependencies, by streaming experts from disk.

  19. GLM 5.2 is (nearly) as accurate as a human book-keeper at less than 1% of the cost We evaluated the performance of GLM 5.2, an open weights AI model, on the task of quarterly value-added tax (VAT) return preparation for a small UK business…

  20. For the Background: I'm building a SaaS in Scala/Play + React. I use AI heavily for coding, not just for suggestions but for full feature implementation, PR reviews, and architecture discussions.

  21. A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten from models like Claude or GPT, and this really surprised me.

  22. We analyze how four forces restructure the AI industry over 2026-2030: the DRAM/HBM price surge, frontier-capable open-weight models (GLM-5.2), rapid inference-efficiency gains (near-Shannon-limit KV-cache compression, lightweight local ru…

  23. VisionBridge Give text-only LLMs vision through a tiny OpenAI-compatible proxy. VisionBridge sits between your chat UI and your models.

  24. mulot [-4285F4?logo=googlechrome&logoColor=white)]() Agentic AI web pentester that drives a browser. An open-weights LLM (GLM-5.2, Gemma or Qwen) drives a real headless Chromium through a Burp-style toolkit and works a target the way a hum…

  25. Hiya! So I've been playing around with having Claude make videos for a bit now even had some success posting the results to TikTok (and setup a whole pipeline so Claude can generate and post autonomously).

  26. GLM 5.2 and the coming AI margin collapse (part 1) This is a two part series focusing on what I believe is perhaps the least understood upcoming shift in AI economics. If you've enjoyed this and want to be notified about the second post, p…

← all threads