Synthszr Charts — die großen AI-Marken im Wettkampf ums Podium
synthszr charts

Unknown · v2 · 4× · naposledy 20. 8. 2026

54
Momentum

DFlash 2 is a "drafter" for speculative decoding of large language models, released by Inco AI (building on the original DFlash technique developed at Z Lab). It is not physical hardware, but an Open Source model/software method that predicts the output of a target LLM (e.g., Qwen3.8-27B, Meta Muse Glimmer) in parallel blocks, thereby increasing Inference speed without changing output quality. The models are available open source under Apache 2.0 license on Hugging Face and run on common Inference engines such as SGLang, vLLM, llama.cpp, and oMLX/MLX on Apple Silicon.

Vývoj momenta
22.05.20.08.

Vlastnosti

LicenseApache 2.0
PlatformRuns in SGLang, vLLM, llama.cpp, and oMLX/MLX (Apple Silicon) from day one
PriceFree (open-source model, no purchase price)
Release DateAugust 18, 2026
MemoryModel size 2B parameters (BF16 ~3.86 GB; Q4_K_M ~1.14 GB; Q8_0 ~2.06 GB)
AvailabilityPublicly available on Hugging Face (incoai/z-lab collections) and GitHub (z-lab/dflash)

Další produkty v této kategorii: AI inferenční hardware

Zdroje (4)

Subscribe free. Unsubscribe the second it sucks.

High-signal news across AI, business, UX, and tech. Every morning.