The announcement-to-weights gap is the most revealing number in open-weight AI that nobody tracks. Every 2026 model launch produces two dates — the day the vendor said “open weights” and the day a checkpoint actually appeared on Hugging Face — and the distance between them varies from zero to open-ended, sometimes inside the same lab’s own release history.
The spread is the story. Moonshot AI shipped Kimi K2.5’s weights the same day it announced them in January, then ran a ten-to-eleven-day gap on Kimi K3 in July. Z.ai put GLM-5.1 on Hugging Face the day it was announced in April, then announced GLM-5.3 on August 14 with no weights at all — and as of this ledger’s August 23 dateline, they still have not shipped. NVIDIA announced a 550B-parameter model from a keynote stage and closed the gap in three days.
This ledger lines up every notable 2026 open-weight release — announcement date, weights-live date, and the gap in days — with a confidence flag on every row. Where two credible sources disagree on a date, the row shows a range or “not established” instead of a number we cannot defend. That restraint is deliberate: a ledger you can cite is worth more than a ledger that looks complete.
- 01Seven 2026 releases shipped weights on announcement day.Kimi K2.5, Gemma 4, GLM-5.1, both Tencent Hy3 drops, DeepSeek V4 Preview, and Nemotron 3.5 Lightning all show public weights on the day the vendor announced them. Gemma 4 is the one medium-confidence row in that set — its exact commit hour is unconfirmed.
- 02The longest confirmed gap is Kimi K3’s 10–11 days.Moonshot announced K3 at WAIC on July 16; the weight dump landed July 26 or 27 depending on the source. Alibaba’s Qwen3.8-Max checkpoint ran a comparable 9–10 days behind its August 3 announcement.
- 03GLM-5.3 is the ledger’s open row.Z.ai announced GLM-5.3 on August 14 as API and Coding Plan only, holding weights for a safety evaluation. As of August 23 they have not shipped — the row reads ‘announced, weights pending’ with no gap computed.
- 04Announced-as-open and downloadable are different events.Several outlets called Kimi K3 ‘open-weight’ days before anyone could download it. This ledger only closes a row when weights are verifiably live on a public repo — a licence commitment is not a release.
- 05Where sources disagree, the ledger keeps the range.GLM-5, GLM-5.2, DiffusionGemma, and MiniMax M3 all have unconfirmed or disputed weights-live dates. Those rows show ranges or ‘not established’ — the confidence column is the citation value, not a weakness.
01 — MethodAnnounced is not downloadable.
The ledger measures exactly one thing: the number of days between a vendor’s own public announcement of an open-weight model and the first date the weights were verifiably available on Hugging Face or an equivalent public repository. The “announced” column comes from the vendor’s blog post, keynote, or launch coverage dated to the announcement itself. The “weights live” column requires either a model-card commit date or at least one independent press report explicitly stating the weights were downloadable.
That two-source discipline matters because the press routinely conflates the two events. Coverage of Kimi K3 described it as “open-weight” from July 17 — when what existed on July 17 was a licence commitment, not a download link. The weights arrived roughly ten days later. A ledger that took launch-day language at face value would erase the very gap it exists to measure.
Three honesty rules govern every row. First, where two credible sources give dates a day or two apart, the row shows the range — Kimi K3 reads “10–11 days”, not a single integer. Second, where the weights-live date could not be confirmed at all, the gap column reads “not established” rather than an estimate. Third, licence terms are a secondary column only — our licence audit of 30 open-weight models owns that story, reading each licence text directly; this ledger just records the tag and links out.
One methodological caveat cuts the other way: sometimes weights are effectively public before the announcement. Meituan’s LongCat-2.0 circulated on OpenRouter as the anonymous “Owl Alpha” listing before the company’s formal reveal — so for some users, first access predated the announcement date entirely. The ledger measures the vendor’s public commitment against the vendor’s public delivery; stealth aliases are noted, not counted.
02 — The LedgerTwenty rows, one axis: days from announcement to weights.
The table below is the primary asset. Rows are grouped by outcome — same-day ships, multi-day gaps, and rows that are still open — and each carries a confidence flag. The licence column records the tag only; for what those licences actually permit, see the full licence audit. Dates are ISO-format, 2026 unless noted.
| Model (vendor) | Announced | Weights live | Gap (days) | Licence tag | Confidence |
|---|---|---|---|---|---|
| Same-day and near-same-day ships | |||||
| Kimi K2.5 (Moonshot AI) | 2026-01-27 | 2026-01-27 | 0 | Modified MIT | High — two sources |
| GLM-5 (Z.ai) | Feb window (~02-11 to 02-17) | same window; exact day unconfirmed | not established | MIT | Low — Feb sequencing unclear |
| Gemma 4 (Google DeepMind) | 2026-04-02 | 2026-04-02 | 0 | Apache 2.0 | Medium — commit hour unconfirmed |
| GLM-5.1 (Z.ai) | 2026-04-08 | 2026-04-08 | 0 | MIT | High |
| Hy3 preview (Tencent) | 2026-04-23 | 2026-04-23 | 0 | Custom — region-restricted | High |
| DeepSeek V4 Preview (Pro + Flash) (DeepSeek) | 2026-04-24 | 2026-04-24 | 0 | MIT | High — three sources |
| DiffusionGemma (Google DeepMind) | 2026-06-10 (reported) | reported same day | reported ~0; unverified | Apache 2.0 | Medium — date not yet checked against Google’s own blog |
| GLM-5.2 (Z.ai) | June window (~06-13 to 06-17) | same window; exact day unconfirmed | not established | MIT | Low — no authoritative day |
| Hy3 full release (Tencent) | 2026-07-06 | 2026-07-06 | 0 | Apache 2.0 | High |
| Nemotron 3.5 Lightning (NVIDIA) | 2026-08-11 | 2026-08-11 | 0 | see repo | High — NVIDIA blog + two press |
| Multi-day gaps | |||||
| Nemotron 3 Ultra (NVIDIA) | 2026-06-01 (Computex keynote) | 2026-06-04 | 3 | see repo | High — NVIDIA blog + two press |
| MiniMax M3 (MiniMax) | ~2026-06-01 | disputed — sources conflict | not established | see repo | Low — weights date unreconciled |
| LongCat-2.0 (Meituan) | 2026-06-30 | 2026-07-04 | 4 | MIT | High — two sources |
| Kimi K3 (Moonshot AI) | 2026-07-16 (WAIC Shanghai) | 2026-07-26 or 07-27 (disputed) | 10–11 | see repo | High on announcement; medium on ship day |
| Qwen3.8 preview (Alibaba) | 2026-07-19 | n/a — API-only preview | n/a | none disclosed | High |
| Qwen3.8-Max (2.4T checkpoint) (Alibaba) | 2026-08-03 | 2026-08-12 to 08-13 | ~9–10 | Custom — MAU/revenue thresholds | Medium — two-day window |
| Qwen3.8-27B (Alibaba) | 2026-08-03 (same announcement) | 2026-08-13 to 08-14 | ~10–11 | Apache 2.0 | Medium — commit date unconfirmed |
| Open rows — pending, teased, or never shipped | |||||
| GLM-5.3 (Z.ai) | 2026-08-14 | pending as of 2026-08-23 | no gap — weights pending | n/a — API + Coding Plan only | High that weights have not shipped |
| Unnamed MoE family (Mistral AI) | teased Jul 2026 — partner early access | not public as of 2026-08-23 | no gap — nothing to measure | n/a | Low — no name, no repo |
| Llama 5 (Meta) | no announcement, no release date | never shipped | n/a — no release to measure | n/a | High that no weights exist |
Two reading notes. The “see repo” licence cells mark rows where this pass recorded the release but not the licence tag — the licence audit is the authority there. And the July rows are ledger entries only: the fuller story of that month’s release cluster lives in July’s open-weight momentum tracker, which this page deliberately does not retell.
03 — Same-Day CohortSeven releases where announcement day was weights day.
Seven rows ship at gap zero, and they span the whole field: a 1T-parameter multimodal model (Kimi K2.5, January 27), a 754B agentic MoE (GLM-5.1, April 8), Google’s Gemma 4 family (April 2), both of Tencent’s Hy3 drops (the region-restricted April preview and the Apache 2.0 full release on July 6), DeepSeek’s V4 Preview pair (April 24), and NVIDIA’s Nemotron 3.5 Lightning (August 11), which hit Hugging Face, ModelScope, and OpenRouter the day it was announced.
Same-day shipping is not a coincidence of logistics — it is an operational stance. A vendor that uploads weights before the announcement is treating the download link as part of the announcement, which means every launch-day claim can be checked within hours. DeepSeek went further still, shipping V4-Pro and V4-Flash simultaneously across weights, API, and its chat surface — the pattern we documented in our V4 Preview launch coverage.
0-day releases
Kimi K2.5, Gemma 4, GLM-5.1, Hy3 preview, DeepSeek V4 Preview, Hy3 full, and Nemotron 3.5 Lightning — same-day weights across six different labs.
Kimi K3
Announced at WAIC Shanghai July 16; the full weight dump landed July 26 or 27 — the two dates are disputed by one day across sources, so the ledger keeps the range.
Pending, teased, never shipped
GLM-5.3 announced with weights pending; Mistral’s unnamed MoE family teased with no public artifact; Llama 5 never formally announced, no release date ever set.
Two rows sit near this cohort but do not earn a hard zero. DiffusionGemma is widely reported as a June 10 same-day drop with day-zero inference support across vLLM, Transformers, MLX, and SGLang — but this pass could not verify the date against Google’s own blog, so the row reads “reported ~0; unverified”. GLM-5 and GLM-5.2 both appear to follow Z.ai’s same-day pattern, but neither February nor June produced a single authoritative day this pass could pin down, so both gaps stay “not established” rather than being forced to zero.
04 — Multi-Day CohortThree days to eleven: when the download link lags the press release.
The multi-day cohort splits into two very different patterns. The short gaps are event-calendar artifacts: NVIDIA announced Nemotron 3 Ultra (550B total / 55B active) from the Computex keynote stage on June 1 and had weights live on June 4 — a three-day gap where the announcement date was fixed by a keynote schedule, not by upload readiness. Meituan’s LongCat-2.0 (1.6T total / ~48B active) ran four days from its June 30 announcement to a July 4 Hugging Face upload.
The long gaps are staged releases. Moonshot unveiled Kimi K3 — 2.8T total parameters, 16 of 896 experts active per token, roughly a 50B-parameter active compute cost — at WAIC on July 16, and the weight dump landed ten to eleven days later. Launch coverage at the time reported K3 placing 4th on the Artificial Analysis Intelligence Index and 1st on Arena.ai’s Frontend Code Arena — reported-at-announcement figures, noted here only as colour. The same July window also saw the White House publicly accuse Moonshot of distilling Anthropic’s Fable model and Treasury Secretary Scott Bessent threaten sanctions — a one-line reminder that weight-release timing on Chinese models now carries political freight beyond logistics.
“The 2.8-trillion-parameter figure describes capacity, not compute... The actual compute cost per token resembles a roughly 50-billion-parameter dense model, not a 2.8-trillion-parameter one.”— Tech Times, reporting on Kimi K3’s architecture, July 25, 2026
Alibaba ran the most elaborate staging of the year. It previewed “Qwen3.8” on July 19 — API-only, no benchmark table, no repo, no licence file — with the release language “Qwen3.8 is launching and going open-weight soon.” The full Qwen3.8-Max announcement followed on August 3, and the 2.4T-parameter open checkpoint went live on Hugging Face around August 12–13: roughly 9–10 days after the full announcement, 24–25 days after the preview. A smaller Qwen3.8-27B under Apache 2.0 followed around August 13–14.
The staging carried a substantive catch: the open checkpoint is narrower than the hosted product — text-only, no vision, and without the API version’s native 1M-token context. That is exactly the kind of announcement-versus-artifact drift this ledger exists to surface: the thing announced on August 3 and the thing downloadable on August 13 are not the same thing.
“This also marks the first time we will open-source the weights of a Qwen-Max-class model.”— Alibaba Qwen team, Qwen3.8-Max release announcement, August 3, 2026
Announcement → weights-live gap · 2026 rows
Sources: vendor blogs and press reports cited in the ledger. Ranges shown where sources disagree; unconfirmed rows (GLM-5, GLM-5.2, DiffusionGemma, MiniMax M3) and pending rows are excluded from this chart.One row belongs in this cohort but cannot be plotted. MiniMax announced its M3 flagship around June 1; one source states weights were live on Hugging Face by June 7, while another framed a similar-length window as merely expected. Those two claims were never independently reconciled, so M3’s gap stays “not established” — and its 59.0% SWE-Bench Pro score is a vendor-stated figure, for the record, not an independent benchmark.
05 — Open RowsThree ways a gap never closes.
The most instructive rows in the ledger have no gap number at all — and they got there by three entirely different mechanisms. One is a deliberate, safety-cited hold. One is a tease that never produced a public artifact. And one is a release that was never formally announced in the first place.
GLM-5.3
Z.ai shipped GLM-5.3 as API and GLM Coding Plan access only, saying weights would follow once a safety evaluation and hardening pass finished — citing cybersecurity capability that grew faster during training than anticipated (CyberGym 84.5%, vendor-stated). As of August 23, no weights.
Mistral MoE family
CEO Arthur Mensch confirmed a new ‘fat but sparse’ Mixture-of-Experts open-weight family entering early access for select partners in July. No model name, benchmark, or Hugging Face repo has surfaced since — there is nothing to measure. Mistral’s last confirmed open flagship remains Mistral Large 3, from December 2025.
Llama 5
Meta has not released a Llama 5, and no formal announcement with a release date was ever made. Its flagship open-weight model remains Llama 4 Maverick from April 2025, and the previewed Behemoth teacher model never publicly shipped.
The GLM-5.3 row is the cleanest anchor for the whole ledger’s thesis, precisely because of Z.ai’s own history: GLM-5.1 shipped weights same-day in April, and GLM-5 and GLM-5.2 both appear to follow the same pattern. GLM-5.3 broke it — the announcement came with API access and a stated intent to publish weights after safety evaluation, and as of this dateline the row still reads “pending”. Same lab, same model family, a completely different release posture. When the weights do land, this row gets its number; until then, the honest entry is the open one.
The Llama 5 row deserves its own footnote on reporting hygiene. The neutral fact is simple: no Llama 5 weights have ever shipped and no release date has ever been announced, a trajectory we traced in our H1 2026 open-weight retrospective. Yet the reporting around it is where this space gets murky.
06 — AnalysisThe gap is a governance signal, not a logistics footnote.
Read across all twenty rows, the pattern is too consistent to be noise. Labs that treat open weights as the product — DeepSeek, early-2026 Z.ai, Google’s Gemma line, Tencent by July — ship at gap zero, because the upload is finished before the press release goes out. Labs that announce from an event stage run short, calendar-driven gaps: NVIDIA’s three days from a Computex keynote is the disciplined version of that pattern. And labs that stage releases — announcing first, uploading later — run gaps of nine days and up, during which the announcement does marketing work that no downloadable artifact can yet verify.
Looking forward, we expect the GLM-5.3 pattern to become more common, not less. As open models post stronger agentic and cybersecurity results, deliberate safety-review holds between announcement and weights are a rational vendor response — which means the gap will increasingly encode a lab’s governance posture rather than its upload bandwidth. That makes this ledger’s axis more informative over time: a widening gap on a vendor’s next release is signal, and an announcement with no dated weights commitment is a row you should treat as open until proven closed.
Same-day drop
The launch-day claims are checkable immediately — benchmarks, licence text, and checkpoint contents can all be verified against the actual artifact within hours. The announcement and the release are the same event.
Event-calendar lag
Keynote announcements fix the marketing date before the upload date. NVIDIA’s Nemotron 3 Ultra (3 days) and Meituan’s LongCat-2.0 (4 days) both closed fast. Normal logistics — but hold evaluation until the weights land.
Staged release
Kimi K3 and Qwen3.8-Max ran 9–11 day gaps, and Qwen’s shipped checkpoint was narrower than the hosted API version — text-only, no native 1M context. Check checkpoint parity before assuming the download matches the announcement.
Hold, tease, or silence
GLM-5.3’s safety hold, Mistral’s unnamed tease, and Llama 5’s non-announcement (no release date ever named) are three different mechanisms with one reader posture: no dated public artifact means no row to build a plan on.
For teams, the operational translation is straightforward: gate any self-hosting or fine-tuning decision on the weights-live date, not the announcement date — and on the shipped checkpoint’s actual capabilities, not the launch post’s. Our guide to which open-weight coding models you can actually self-host picks up exactly where this ledger stops. And if you are deciding whether an open-weight model belongs in your production stack at all, that evaluation is the first step of our AI transformation engagements.
07 — ConclusionA ledger that gets more useful every time a model ships.
The distance between a press release and a download link is now a number worth tracking.
Twenty rows into 2026, the announcement-to-weights gap runs from zero to open-ended, and the spread is widening at both ends. Seven releases shipped weights the day they were announced. The staged releases ran gaps of nine days and up. One — GLM-5.3 — is announced, API-live, and still weightless on this dateline, held back by its own vendor’s safety review.
The rows without numbers are the point, not the flaw. Where sources dispute a date by a day, the ledger keeps the range; where the weights date could not be confirmed, the gap reads “not established”; where nothing shipped, nothing is measured. A gap table that invented integers for GLM-5.2 or MiniMax M3 would be more complete and less true — and in a space where at least one aggregator has published an entirely fictional Llama 5 launch, complete with an invented release date, true is the scarcer commodity.
This is a maintained ledger. As pending rows resolve and new models ship, rows get appended and ranges get tightened to single dates when primary sources allow it. The announcement is marketing; the weights are the release. This page will keep measuring the distance between the two.