Topic

#open-weights

18 articles tagged open-weights. Browse the full set below, or see all topics.

Tagged "open-weights"

Cross-cutting reads on this topic

18 articles
DeepSeek published DeepSeek-V4-Flash-Vision-Exp's weights under MIT on August 31, ten days after the API-only launch, with a reference implementation attached.
#deepseek#open-weights+5 more
2026-08-31
Read Article
Z.ai published GLM-5.3's 753B weights on August 28 under a bespoke licence, two days after shipping GLM-5.3-Flash under plain MIT. The clause that differs.
#glm-5-3#z-ai+5 more
2026-08-28
Read Article
In the 2025-26 record we checked, two labs gave a public reason before holding back open weights. What a stated delay tells a buyer, and what it does not.
#open-weights#model-releases+5 more
2026-08-28
Read Article
fal says H3 Max renders five seconds of video in under three. The claim is the vendor's own, and the tracker it cites backs the quality half only in part.
#fal#minimax h3+5 more
2026-08-27
Read Article
Alibaba open-weighted Qwen3.8-Flash-Next on August 26. The hosted Qwen3.8-Flash is a documented, different artifact — licence, context and tools all diverge.
#qwen3-8-flash-next#qwen+4 more
2026-08-26
Read Article
DeepSeek's changelog now carries a GA entry for V4-Pro across app, web, and API, MIT-licensed 0813 weights on Hugging Face, and pricing changing August 16.
#deepseek#deepseek-v4-pro+5 more
2026-08-15
Read Article
GLM-5.3 reuses GLM-5.2's base model, so Z.ai says every gain is post-training. Terminal-Bench 3.0 jumped 4.6 to 28.3, and weights are promised later.
#glm-5-3#z-ai+5 more
2026-08-14
Read Article
Mistral's Shieldstral reads your safety policy as prompt text at inference time. A 3B Apache-2.0 guard model with a mixed, not one-sided, benchmark record.
#mistral#shieldstral+5 more
2026-08-10
Read Article
Alibaba promised Qwen3.8-Max and a 27B sibling for the week of 10 August. Updated: Max weights landed 8 August under a bespoke licence. Run these checks.
#qwen#open-weights+5 more
2026-08-07
Read Article
At the time of writing there is no Wan 3.0: no Alibaba announcement, no weights, no listing. July really shipped one open-weight model and two research papers.
#wan-3-0#alibaba+5 more
2026-08-06
Read Article
Thinking Machines' 276B/12B Inkling-Small leads its 975B parent on the vendor's own coding and reasoning table, then loses badly on factuality.
#open-weights#thinking-machines+5 more
2026-08-01
Read Article
OpenAI, Anthropic and Google skipped NVIDIA's Open Secure AI Alliance a month after joining Akrites. That selectivity, not the absence itself, is the story.
#ai-governance#agent-security+5 more
2026-08-01
Read Article
MiniMax launched H3, a 2K video model with native synced audio at $0.13 per second. Weights are promised within days, not published yet.
#MiniMax#AI Video+4 more
2026-07-31
Read Article
Kimi K3's weights are open, but the licence is bespoke: a $20M revenue trigger for resellers, UI attribution rules, and a 1.56TB download to plan for.
#Kimi K3#Open Weights+4 more
2026-07-27
Read Article
Kimi K3 is 1.56TB of weights and needs 64+ accelerators to serve. Here is the real cost maths of self-hosting a trillion-parameter model in 2026.
#Self-Hosting AI#GPU Infrastructure+4 more
2026-07-27
Read Article
Tencent's Hy3 shipped Apache 2.0 at a sub-300GB FP8 footprint. A routing guide: send agentic tool-use to Hy3 and repo-scale coding to a frontier model.
#hunyuan-hy3#open-weights+5 more
2026-07-07
Read Article
Gemma 4 12B processes text, image, audio, and video with no separate encoders, fitting in ~7GB at 4-bit. A guide to running private multimodal agents locally.
#gemma-4#local-ai+6 more
2026-06-10
Read Article
NVIDIA shipped Nemotron 3 Ultra, a 550B open MoE reasoning model with weights, data and recipes under a permissive license. It runs fast but trails Kimi K2.6.
#nvidia#nemotron+6 more
2026-06-05
Read Article