Pelican on a Bicycle: Claude Fable 5 vs. GPT-5.5 Pro vs. Gemini 3.1 Pro (www.promptfrenzy.com via hn)
model roundup
Gemini 3.1
-
Pelican on a Bicycle: Claude Fable 5 vs GPT-5.5 Pro vs Gemini 3.1 Pro We asked the top frontier AI models — launch-day Claude Fable 5, GPT-5.5 Pro and Gemini 3.1 Pro — to draw a pelican riding a bicycle as SVG code. Same prompt, one shot,…
-
Fable 5 below even Gemini 3.1 on Livebench (www.reddit.comhttps)
Is this benchmark broken, or is Anthropic benchmaxing? LiveBench
-
Fable 5 benchmark with remotion video (www.reddit.comhttps)
Overall an improvement over Opus 4.8, but I'd still say Gemini 3.1 Pro has more of an artistic vision even tho it fails tool calls and writes buggy code sometimes. Ik almost everyone is interested just in the SWE stuff, but this has been a…
-
Gemma 4 26B A4B IT QAT Comparison (www.reddit.com via reddit)
Hopefully this isn't too low effort of a post. I just finished the benchmarks and I figured I'd post them online because they certainly were insightful for me.
-
Intresting! Gemini 3.1 has strongest world knowledge but still choose to be lazy (www.reddit.com via reddit)
could not extract summary
-
I Compared the Top AI Models of 2026 — The Results Were More Nuanced Than Expected (www.reddit.com via reddit)
Over the last few weeks I've been comparing the latest frontier AI models, including Claude Opus 4.8, GPT-5.5, Gemini 3.1 Pro, Grok 4.3, Perplexity AI and DeepSeek V4-Pro. Instead of focusing only on benchmark scores, I looked at: Real-wor…
-
Ask HN: Is it feasible to run a model on device for complete privacy? (news.ycombinator.com)
Tried Gemma, Qwen and a few others. Need vision and larger context windows for an application I am working on.
-
Title