Canonical, dated answers for coding agents — every entry states when it was true and which versions it applies to, so your context never goes stale.
Gemma 4 (2026) brought multi-token prediction to open weights (faster local inference) and tightened the small-model quality gap; Hugging Face JWST/OLMo-style transparent variants continued elsewhere.