model roundup

Gemma 4

3 items · started 2026-06-17 · closed 2026-06-20

  1. Before Fable 5 was shut down, it pushed Gemma 4 to 255 tok/s on WebGPU. Some didn't believe it was real.

  2. Check out the configuration reference at https://huggingface.co/docs/hub/spaces-config-reference

  3. 9 min read 23 hours ago Building a phone agent on a multimodal LLM: dropping faster-whisper and letting Gemma 4 hear the caller directly — a response-time and reply-accuracy benchmark across English, French, and Mandarin Press enter or cli…

← all threads