Tagged "deepseek"
Cross-cutting reads on this topic
DeepSeek V4.1 Flash adds vision and lower prices. Compare 19 official benchmarks, agent frameworks and API costs; V4 Pro now continues past September 14.
#DeepSeek#Model Releases+3 more
2026-09-09
Read Article
DeepSeek published DeepSeek-V4-Flash-Vision-Exp's weights under MIT on August 31, ten days after the API-only launch, with a reference implementation attached.
#deepseek#open-weights+5 more
2026-08-31
Read Article
DeepSeek shipped a separate experimental V4 Flash vision model on August 21 at text prices, on a peak and off-peak clock that doubles the rate.
#deepseek#vision models+5 more
2026-08-21
Read Article
DeepSeek's vision model bills each image at up to 384 tokens with no vision surcharge, which turns a 20,000-SKU catalogue QA pass into arithmetic.
#ecommerce#product catalogue qa+5 more
2026-08-21
Read Article
Three Chinese labs shipped major releases and their licence stories diverge. A buyer's scoreboard on weights, pricing schedules, and benchmark claims.
#deepseek#glm-5-3+5 more
2026-08-15
Read Article
Time-of-day LLM pricing is arriving, and a cheaper window is not the same as a cheaper bill. How the announced schedules and peak hours actually work.
#llm-pricing#off-peak-pricing+5 more
2026-08-15
Read Article
DeepSeek open-sourced its dsh agent harness under MIT: a developer preview at 0.1.0-rc.5, four runtime modes, and every part of the product a plugin.
#deepseek#deepseek-harness+5 more
2026-08-15
Read Article
DeepSeek's changelog now carries a GA entry for V4-Pro across app, web, and API, MIT-licensed 0813 weights on Hugging Face, and pricing changing August 16.
#deepseek#deepseek-v4-pro+5 more
2026-08-15
Read Article
On August 12, DeepSeek's pricing page listed V4-Pro-0813 with no changelog entry and a homepage saying nothing had changed. GA was announced a day later.
#deepseek#model-versioning+5 more
2026-08-12
Read Article
GPT-5.6 Luna fell 80% and Terra 20% on July 30. Sonnet 5's scheduled $3/$15 increase was cancelled. OpenRouter list prices mix standard with batch rates.
#ai-api-pricing#gpt-5-6+5 more
2026-08-05
Read Article
One week repriced the cheap end of the model market. A surface-labelled table of what Luna, Qwen3.7 Flash and DeepSeek V4 Flash now cost, and what that buys.
#llm-pricing#cost-optimization+5 more
2026-08-01
Read Article
Worked cost math for classification, extraction, dedup and translation at million-row scale, plus the benchmark evidence on where cheap models should not go.
#deepseek#bulk-processing+5 more
2026-08-01
Read Article
DeepSeek V4 Flash adds OpenAI's Responses API alongside the Anthropic format it has run since 2025, so Codex CLI or Claude Code needs only a base-URL swap.
#DeepSeek#API Design+4 more
2026-07-31
Read Article
DeepSeek V4 Flash exits preview into public beta as the 0731 checkpoint: same 284B architecture, vendor-stated agent benchmarks, no weights posted yet.
#DeepSeek#Open Models+4 more
2026-07-31
Read Article
DeepSeek aliases retired, MCP capabilities deprecated, Sora's API sunsetting — three unrelated clocks in one month. How to build a deprecation calendar.
#model-deprecation#api-lifecycle+5 more
2026-07-30
Read Article
DeepSeek retires its deepseek-chat and deepseek-reasoner API aliases on July 24, 2026 at 15:59 UTC. The confirmed migration steps, plus the claims to watch.
#deepseek#api-migration+4 more
2026-07-20
Read Article
Hunt.io found Claude Code and DeepSeek-v4-pro wired into a suspected China-linked intrusion. The agent-governance controls that close the gap it exploited.
#ai-agent-governance#claude-code+5 more
2026-07-18
Read Article
OpenAI's GPT-5.6 caching overhaul, plus Anthropic and DeepSeek tiers, reshapes agent cost math. How to design cache-first, model-homogeneous agents.
#AI Development#Prompt Caching+5 more
2026-07-16
Read Article
US frontier-model gating is pushing global builders toward open-weight models China already dominates. Does Washington's lockdown hand Beijing the edge?
#China AI#Open-Source AI+6 more
2026-06-27
Read Article
DeepSeek abandons its no-outside-capital stance in a ~$7.4B maiden round led by Tencent and CATL, valuing it near $59B and reshaping open-weight economics.
#deepseek#ai-funding+6 more
2026-06-03
Read Article
GPU spend, ops headcount, latency, and break-even volume for hosting Llama, Qwen, DeepSeek, and Mistral yourself vs API. With per-token cost curves at 4 scales.
#self-hosting-llm#ai-tco+8 more
2026-04-24
Read Article
Q2 2026 market share report on Chinese AI providers — Qwen, GLM, DeepSeek, Kimi, MiniMax, Baichuan, and Yi. Usage data, licensing, and enterprise adoption.
#chinese-ai-models#qwen+4 more
2026-04-12
Read Article
Anthropic accuses DeepSeek, Moonshot AI, and MiniMax of industrial-scale distillation via 24,000 fake accounts and 16M+ Claude exchanges. Full analysis inside.
#ai-distillation#anthropic+5 more
2026-02-24
Read Article
DeepSeek V4 brings 1 trillion parameters, 1M token context, and Engram O(1) memory. Architecture details, leaked benchmarks, and what it means for developers.
#DeepSeek#DeepSeek V4+6 more
2026-02-14
Read Article
DeepSeek-V3.1 guide preserved as historical 2025 context, with 2026 caveats for newer DeepSeek V3.2 and V4 reasoning models.
#DeepSeek#Open Source AI+5 more
2025-08-22
Read Article