Topic

#AI Models

27 articles tagged AI Models. Browse the full set below, or see all topics.

Tagged "AI Models"

Cross-cutting reads on this topic

27 articles
A 4B model trained for $1,200 cut Postgres query latency 44.7% on a standard benchmark. The trait that made it work, and a routing table for your own tasks.
#AI Models#Fine-Tuning+2 more
2026-09-18
Read Article
Qwen's September 18 model takes audio and video natively with 1M context and, by Qwen's own method, cuts per-hour audio cost 98% versus its predecessor.
#AI Models#Agentic AI+3 more
2026-09-18
Read Article
TypeSafe AI's Jev returns typed values with probabilities, not text, and claims 70 to 500 ms responses. What the vendor figures mean and how to test them.
#AI Models#Agentic AI+2 more
2026-09-16
Read Article
Compare Fugu Max and Ultra v2 pricing, long-context charges and review costs. Use a workload test to decide when additional reasoning justifies its price.
#AI Models#Fugu+3 more
2026-09-12
Read Article
A dated ledger of AI model releases in September 2026, each row verified against the vendor’s announcement, with price, context and what it replaces.
#model-releases#ai-models+4 more
2026-09-02
Read Article
On August 12, DeepSeek's pricing page listed V4-Pro-0813 with no changelog entry and a homepage saying nothing had changed. GA was announced a day later.
#deepseek#model-versioning+5 more
2026-08-12
Read Article
xAI's Grok 4.6 keeps Grok 4.5's headline rates, but the $2/$6 tier only holds below a 200K-token prompt: cross it and the whole request reprices.
#grok-4-6#xai+5 more
2026-08-12
Read Article
Alibaba previewed Qwen3.8-Max at WAIC Shanghai, claiming 2.4 trillion parameters and second only to Fable 5 — yet shipped zero benchmarks to back it.
#qwen#alibaba+5 more
2026-07-20
Read Article
Twelve AI models launched in a single week of March 2026 from OpenAI, Google, Mistral, xAI, and more. Developer guide to capabilities, pricing, and selection.
#ai-models#march-2026+4 more
2026-03-15
Read Article
xAI releases Grok 4.20 with 2M token context, lowest hallucination rate at 78%, and 60% lower pricing. Three API variants for reasoning and multi-agent use.
#grok-4-20#xai+4 more
2026-03-10
Read Article
GPT-5.4 ships three variants: Standard, Thinking, and Pro. Native computer use, 1M context, tool search, and 33% fewer factual errors. Complete guide.
#gpt-5-4#openai+5 more
2026-03-06
Read Article
Sarvam AI releases 105B and 30B open-source reasoning models under Apache 2.0, trained entirely in India. 128K context window powers the Indus AI assistant.
#sarvam-ai#open-source-ai+4 more
2026-03-06
Read Article
OpenAI teased GPT-5.4 the same day as GPT-5.3 Instant launch. Rumored 2M context window, enhanced reasoning, and what it means for the AI model roadmap.
#gpt-5-4#openai+4 more
2026-03-05
Read Article
OpenAI releases GPT-5.3 Instant with 26.8% fewer hallucinations, 400K context, and anti-cringe tone overhaul. Complete benchmarks, pricing, and migration guide.
#gpt-5-3-instant#openai+4 more
2026-03-03
Read Article
Google launches Gemini 3.1 Flash-Lite at $0.25 per million input tokens. 2.5x faster, tops 6 benchmarks. Complete pricing and performance comparison guide.
#gemini-flash-lite#google-ai+4 more
2026-03-03
Read Article
OpenAI's GPT-5.3-Codex brings 25% faster inference and major Terminal-Bench and OSWorld gains. Full benchmarks, access details, and migration guide.
#GPT-5.3-Codex#OpenAI+3 more
2026-02-05
Read Article
Claude Opus 4.6 brings 1M token context, adaptive thinking, and 128K output. Complete guide to benchmarks, pricing, API changes, and enterprise features.
#Claude Opus 4.6#Anthropic+3 more
2026-02-05
Read Article
OpenAI launches GPT-5.2 and GPT-5.2-Codex optimized for agentic coding as GPT-4o retires February 13, 2026. Complete migration and implementation guide.
#GPT-5.2#OpenAI+3 more
2026-01-31
Read Article
Master Grok 4.1: EQ-Bench #1 ranking, 65% hallucination reduction, Fast API access, xAI benchmarks, and comparison with GPT-5.2 and Claude Opus 4.5.
#Grok#xAI+5 more
2025-12-17
Read Article
Master GPT-5.2 with Instant/Thinking/Pro tiers. 38% fewer errors, 70.9% expert-level accuracy. Complete guide with benchmarks and integration.
#GPT-5.2#OpenAI+5 more
2025-12-11
Read Article
Master GPT-5.1 Instant and Thinking models. 8 personalities, 2-3x faster. Complete guide with API and ChatGPT integration.
#GPT-5.1#OpenAI+5 more
2025-11-12
Read Article
Deploy GLM 4.6 with current Z.ai, OpenRouter, vLLM, and SGLang guidance covering endpoints, pricing, MIT licensing, model IDs, and production caveats.
#GLM 4.6#Zhipu AI+6 more
2025-10-14
Read Article
Qwen3 model-family guide refreshed with 2026 caveats for Qwen3.6-Max-Preview, Coder, Thinking models, and deployment strategies.
#AI Models#Qwen+6 more
2025-09-08
Read Article
Historical Kimi K2-0905 release guide refreshed for 2026 context, with agentic coding, MoE architecture, deployment, and newer Kimi caveats.
#AI Models#Open Source+3 more
2025-09-05
Read Article
DeepSeek-V3.1 guide preserved as historical 2025 context, with 2026 caveats for newer DeepSeek V3.2 and V4 reasoning models.
#DeepSeek#Open Source AI+5 more
2025-08-22
Read Article
Historical GPT-5 launch guide refreshed with 2026 caveats for GPT-5.4 and GPT-5.5, capabilities, pricing, benchmarks, and API access.
#GPT-5#OpenAI+5 more
2025-08-07
Read Article
Gemini 2.5 Flash, Pro, and Deep Think guide refreshed with 2026 caveats for Gemini 3.1, pricing, benchmarks, and model selection.
#Gemini AI#Google AI+5 more
2025-08-04
Read Article