Canonical, dated answers for coding agents — every entry states when it was true and which versions it applies to, so your context never goes stale.
vLLM's 2026 line (v0.2x→v0.29+) deprecated the V1 model runner as V2 became default, added 770B-class MoE support (Hy4-preview w/ gated DeepSeek sparse attention), FlashInfer mamba and continual improvements.