
DeepSeek's new 304B agentic model now runs on a single 128GB workstation
Salvatore Sanfilippo repacked DeepSeek V4 Flash into a lossless MXFP4 GGUF that streams from SSD at over 20 tokens a second. The hardware bill, and where hosted still wins.

Salvatore Sanfilippo repacked DeepSeek V4 Flash into a lossless MXFP4 GGUF that streams from SSD at over 20 tokens a second. The hardware bill, and where hosted still wins.

A popular Hacker News how-to walked through a fully local coding agent on Apple Silicon. Here's the realistic 2026 stack: runner, model, and harness.

Cyera disclosed CVE-2026-7482 on May 1, a CVSS 9.1 unauthenticated heap read in Ollama. Three API calls dump prompts, env vars, and API keys from any open instance.

A leaked Geekbench listing puts AMD's Ryzen AI Max+ 495 on a 192GB platform with a Radeon 8065S iGPU. The Strix Halo chip it replaces capped at 128GB.

Apple quietly pulled the 256GB Mac mini from its store on May 1. Tim Cook had warned the day before that demand was outpacing supply for months.

Alibaba's Qwen 3.6-35B-A3B is a 35B-param mixture-of-experts with only 3B active. Apache 2.0, runs on consumer GPUs, and it's already winning real tasks.