DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395 (www.lucebox.com via hn)
model roundup
DeepSeek 4
-
July 2026 DeepSeek V4 Flash: 284B model, up to 32 tok/s on AMD Ryzen AI MAX+ 395 AMD-Powered Lucebox runs the full DeepSeek V4 Flash target locally on AMD Ryzen AI MAX+ 395 with 128 GB unified memory: up to 32.0 tok/s decode and roughly 25…
-
I Sat on an Idea for 7 Years. AI Helped Me File for a Patent in 2 Weeks. (pablooliva.de via reddit)
I ran a side-by-side on a real project: Claude Code on a Max plan versus an open-weight agent stack (GLM 5.2 via Hermes Agent, DeepSeek v4 Pro for second opinions), working through a provisional patent application for a product idea I'd sa…
-
We A/B tested Ante's half-size system prompt on deepseek-v4-flash across the full terminal-bench 2.1 suite: no measurable performance change, and among the 69 tasks whose outcome stayed the same, the short-prompt run's median input-token c…
-
We self-host DeepSeek V4 Flash on AWS spot instances (twitter.com via hn)
https://t.co/syBl6fkTDJ Miguel Salinas@VercantezHow we self-host DeepSeek V4 Flash on AWS spot instances4:18 PM · Jul 23, 20269.7KViews243560
-
[AINews] "Laguna S 2.1 Released: Cheaper than Deepseek v4 Flash, Better than V4 Pro" a quiet day lets us highlight a new neolab win. Reignited distillation wars conversation aside, today was more of the same of previous news cycles, which…
-
Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed training, including severe memory pressure, non-overlapped communication overhead, and inefficie…