October 2026 opened without a frontier launch and with a cluster of small, specialised models instead. On October 1 Cloudflare and Amazon each released open-weight decision models, Microsoft AI released two text-to-speech models and a streaming transcription model, and Tavus previewed a full-duplex video model to invited testers. On October 2 Bilibili’s Index team published the preview of a 35B-total, 3B-active translation model. This page records each one on the vendor’s own date and is refreshed in place through the month.
Editorial note: Opened October 3 as an October 1, 2026 dispatch. Rows and listings dated October 2 postdate the dateline and were collected on October 3 with the rest; rows added after October 3 carry the date they were added. Facts are as of October 3 unless a row says otherwise.
- 01Eight models, five vendors, no frontier LLMEvery release in the first three days is a specialist: decision, speech, video or translation. The frontier news is an availability change, not a launch.
- 02Open weights outnumber closedFour of the eight carry Apache 2.0 licences with public weights. Microsoft’s three and Tavus’s one are API or invite only.
- 03Three listings, no vendor pageThree models appeared on OpenRouter on October 1 and 2 with no release page we could locate. They are recorded as listings, not releases.
- 04Argon is still not publicGemini 4 Argon, announced September 30, is reaching trusted cyber defenders first. Its introductory price is published; its general date is not.
On a vendor page, Oct 1–3
Counted by model, so Clef and Clef-flash are two.
Apache 2.0 with public weights
Clef, Clef-flash, Strands Decider 2B and Index-Translate-35B.
On OpenRouter without a release page
Recorded in section 04 and excluded from the counts above.
01 — The ledgerThe October ledger
One row per model, dated from the vendor’s announcement. Where a vendor released two sizes in one post they share a row. The primary column names the page each row was read from; the linked sources are collected in the method section.
| Date | Vendor · model | What it is | Primary source |
|---|---|---|---|
| Oct 1 | Cloudflare · Clef and Clef-flash | Decision models (typed answers with probabilities, no text), 27B and 9B, post-trained from Qwen bases; read text, JSON, images and video | Cloudflare blog, October 1; Hugging Face cards modified October 1 |
| Oct 1 | Amazon Strands Labs · Strands Decider 2B | 2B decision model for local CPU or GPU; weights, training data and scripts released | Strands blog, October 1; Hugging Face card modified October 1 |
| Oct 1 | Microsoft AI · MAI-Transcribe-2-Streaming | Streaming speech-to-text, 60 languages, continuous language detection | Microsoft AI news post, October 1 |
| Oct 1 | Microsoft AI · MAI-Voice-2.1 | Text-to-speech, 23 languages and 26 locales, one voice across languages | Microsoft AI news post, October 1 |
| Oct 1 | Microsoft AI · MAI-Voice-2.1-Flash | Low-latency variant of the above for high-volume use; vendor states 150 ms end to end | Microsoft AI news post, October 1 |
| Oct 1 | Tavus · Griffin-Lite | Research preview of a full-duplex video-to-video conversation model; the full Griffin model is announced for later | Tavus announcement page and Business Wire release, October 1 (not linked) |
| Oct 2 | Bilibili Index team · Index-Translate-35B-A3B-preview | Mixture-of-experts translation model, 35B total and 3B active parameters, 150 text languages; dense 2B and 9B siblings reached Hugging Face on September 28 | Hugging Face card created October 2; Index-Translate project site |
The two decision-model rows are the month’s first story, and they have their own page: open decision models compared puts Clef, Clef-flash and Decider 2B beside TypeSafe’s Jev on licence, size, context and the vendor’s benchmark rows.
02 — The termsPrice, context and licence
| Model | Price | Context | Licence and where it runs |
|---|---|---|---|
| Clef / Clef-flash | $0.24 / $0.09 per million input tokens on Workers AI | 65,536 | Apache 2.0; Workers AI or self-hosted |
| Strands Decider 2B | Free, local | Not stated | Apache 2.0; GitHub and Hugging Face |
| MAI-Transcribe-2-Streaming | $0.54 per hour of audio, introductory through end of 2026 | n/a | Closed; Microsoft AI and Azure |
| MAI-Voice-2.1 | $22 per million characters | n/a | Closed; Microsoft AI, Azure, OpenRouter |
| MAI-Voice-2.1-Flash | $15 per million characters | n/a | Closed; Microsoft AI, Azure, OpenRouter |
| Griffin-Lite | Not published | n/a | Closed; select early testers only |
| Index-Translate-35B-A3B-preview | Free weights | 262,144 in the shipped config; card examples serve 32,768 | Apache 2.0; Hugging Face and ModelScope |
Microsoft’s post attaches two claims to the Flash voice model that are the vendor’s own and unverified here: that inference is 55% faster and roughly 60% cheaper than comparable models, and that it produces 45 seconds of audio at 150 milliseconds end to end. Our September ranking of text-to-speech models is where the two voice models will be scored once listener data exists.
03 — ContextAvailability changes, not releases
The frontier model in the news this week is not in the ledger, because it was announced in September and is not yet generally available. Google announced Gemini 4 Argon on September 30 as rolling out to a set of trusted cyber defenders through its Fairwind programme, with an introductory price of $2 per million input tokens and $10 per million output tokens and cached input at 95% off. The post gives no date for developers, enterprises or consumers. Our Argon page tracks that date.
Three September releases sit just outside this ledger and are not counted here: Together’s Tev1 4B experimental, published on GitHub and Hugging Face on September 23 and listed on OpenRouter on September 30; Inception’s Mercury Decide, listed on OpenRouter on September 30 with no vendor page located; and Bilibili’s dense Index-Translate 2B and 9B, on Hugging Face since September 28, which the October 2 preview extends. The two decision models are covered in decision models that return probabilities, not text. The month before this one is in the September tracker.
04 — CorroborationListings without a vendor page
OpenRouter’s model list added three entries in the window with no vendor release page located by October 3: Unbiased’s Pareto 26.10 Preview on October 1 (1.05M context, $0.80 input and $3.20 output per million tokens), Apodex 1.1 Mini on October 1 (262K, free tier), and inclusionAI’s Ling 3.1 Flash on October 2 (262K, free). A listing time is a listing, not a launch. Each moves into the ledger when a vendor page with a date appears.
05 — AnalysisWhat the first days show
The pattern is the inverse of September’s opening, which brought three frontier models in two days. This month began with models that do one thing: decide, speak, transcribe, translate, or hold a face-to-face conversation. Two of the five vendors are infrastructure companies, Cloudflare and Amazon, releasing weights rather than selling tokens. That is consistent with where decision models sit in an agent: close to the tool call, where a hosted round trip is the cost.
For a buyer, the practical reading is that none of the October rows replaces a model already in production. They add a new slot: a cheap check before an action, a voice that keeps its identity across languages, a translator with open weights. The question to ask of each is not “is it better than the frontier” but “what does it let me stop sending to the frontier”. Our AI transformation work starts from that question.
06 — MethodMethod and as-of date
A census of vendor announcements. No model on this page was run by Digital Applied.
- What was collected
- Every AI model with a vendor-dated release page between October 1 and October 3, 2026, with its price, context window, licence and hosting as stated by the vendor. One row per model, except where a vendor released two sizes in one announcement.
- Sources
- Cloudflare’s Clef announcement and Workers AI model pages; Amazon’s Strands Decider post on the Strands Agents site, dated October 1; Microsoft AI’s transcription and voice post on its news site, dated October 1; the Index-Translate-35B-A3B-preview card; Hugging Face API records for licence tags, base-model tags and creation dates; Tavus’s announcement page and its Business Wire release; the OpenRouter models API for listing times.
- As-of date
- October 3, 2026. The ledger was opened on that date and covers October 1–3. Later rows carry the date they were added.
- Dating rule
- A row’s date is the vendor announcement date. A Hugging Face card created before the announcement does not move the date earlier; an OpenRouter listing never sets a date on its own.
- Exclusions
- Availability changes to models announced earlier, such as Gemini 4 Argon; products, programmes and APIs that are not models; OpenRouter listings with no located vendor page.
- Refresh
- Rows are added as vendors publish through October 31, each marked with the date it was added. Corrections are made in place with a dated note. The month closes with a count of releases by type.
Check the row’s date before you quote it
Every row names the page it was read from and the day it was read. Use the vendor date, not the listing date, and treat the price and latency claims as the vendor’s until a third party measures them.