DeepSeek open-sources inference optimizations with 60–85% faster generation [pdf] from Hacker News on 2026-06-27 09:18 (#76KPF) Comments
DSpark: Speculative decoding accelerates LLM inference [pdf] from Hacker News on 2026-06-27 09:18 (#76KSQ) Comments