Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
DFlash 2 achieved 2.26x speedup on 100 LiveCodeBench coding problems.
Source findingDFlash 2 uses Qwen3.8-27B as the draft model for speculative decoding.
Source findingInco AI shipped DFlash 2 on 2026-08-18 with speculative decoding.
Source findingDFlash 2 integration PR is open for llama.cpp as ggml-org/llama.cpp #27342.
Source findingDFlash 2 quantizer supports Meta Muse Glimmer 30B.
Source findingDFlash 2 achieves 3.1-4.6x throughput on Meta Muse Glimmer 30B.
Source findingDFlash 2 achieves 3.4x throughput on Qwen3.8-27B.
Source findingInco AI released DFlash 2 on 2026-08-18.
Source findingDFlash 2 is implemented in llama.cpp via pull request #27342.
Source findingDFlash 2 quantizer supports Qwen 3.8 27B with 30% claimed faster inference.
Source findingInco AI shipped DFlash 2 on 2026-08-18 with speculative decoding.
Source findingDFlash 2 integration PR is open for llama.cpp as ggml-org/llama.cpp #27342.
Source finding