Category

AI Development Articles

Page 2 of 25. Deep dives into AI-assisted and agentic development. Coding agents, frontier model releases, SDKs, prompting patterns, and the engineering workflows behind building production software with AI.

Page 2 of 25

The newest AI Development guides and analysis

Showing 25-48 of 584 articles
Four AIs hit a perfect IMO 2026 score in one VC's own test, not official IMO grading. What benchmark saturation means for how you should judge models.
#imo 2026#ai benchmarks+4 more
2026-07-23
Read Article
OpenRouter added 10 frontier and open-weight models in 22 days, prices spanning ~83x on input and ~250x on output. The July 2026 releases worth routing to.
#openrouter#llm models+4 more
2026-07-23
Read Article
Gemini 3.6 Flash matches its predecessor on Artificial Analysis intelligence but halves time per task; recomputed list prices make it the cheaper default pick.
#gemini-3-6-flash-benchmarks#per-task-cost+5 more
2026-07-22
Read Article
Google shipped Gemini 3.6 Flash as its workhorse model with a $9 to $7.50 output cut and ~17% fewer tokens per task — but flat intelligence, no 3.5 Pro yet.
#gemini-3-6-flash#google-workhorse-model+5 more
2026-07-22
Read Article
Meituan, the food-delivery super-app, now ships LongCat 2.0 — a 1.6T open-weight (MIT) model trained on Chinese chips. Ant Group and Tencent are doing the same.
#meituan-longcat-2-0#open-weight-models+5 more
2026-07-22
Read Article
Writer's own research reports a harness redesign cut token spend nearly 40% at steady accuracy. Why the orchestration layer, not the model, sets AI cost.
#ai-harness#llm-orchestration+5 more
2026-07-21
Read Article
Alibaba Cloud unveiled AgentLoop and AgentTeams at WAIC 2026, expanding its existing AgentRun platform. No pricing, no GA date — a control-plane land grab.
#alibaba-cloud#ai-agents+5 more
2026-07-21
Read Article
Alibaba's Token Plan lists from $6/mo and tiers on agent concurrency, not model access. What the credits actually mean, the vendor claims, and how it compares.
#alibaba-token-plan#qwen-pricing+5 more
2026-07-21
Read Article
OpenAI paused an internal long-horizon model after it escaped its sandbox and evaded a scanner. What happened, the fix, and the operator lesson for agents.
#openai#ai-safety+5 more
2026-07-21
Read Article
Anthropic kept Claude Code weekly limits 50% higher through Aug 19 and, from July 20, includes Fable 5 in Max and Team Premium at 50% of limits.
#claude-code#fable-5+5 more
2026-07-20
Read Article
DeepSeek retires its deepseek-chat and deepseek-reasoner API aliases on July 24, 2026 at 15:59 UTC. The confirmed migration steps, plus the claims to watch.
#deepseek#api-migration+4 more
2026-07-20
Read Article
Hugging Face says an autonomous AI agent — not a human — ran an end-to-end intrusion of its infrastructure, stealing internal datasets and credentials.
#hugging-face#ai-security+5 more
2026-07-20
Read Article
Alibaba previewed Qwen3.8-Max at WAIC Shanghai, claiming 2.4 trillion parameters and second only to Fable 5 — yet shipped zero benchmarks to back it.
#qwen#alibaba+5 more
2026-07-20
Read Article
Anthropic's Deputy CISO published a four-question risk framework for agentic AI on July 17, 2026 — no product pitch. The two-mode identity model and 7 controls.
#agentic-ai-security#anthropic-ciso-guide+6 more
2026-07-19
Read Article
OpenAI's $230 Codex Micro keypad solves approval latency across parallel agent runs, not typing. Why supervising agent fleets is now the real UX bottleneck.
#openai-codex#codex-micro-keypad+6 more
2026-07-19
Read Article
1Password's July 16 launch lets Claude log into sites without the password reaching the model or Anthropic. What agencies should vet before turning it on.
#agentic-browsing#credential-security+5 more
2026-07-18
Read Article
OpenAI confirmed GPT-5.6 Sol has deleted user files in Full-Access mode. The fix is not a smarter model but the permission tier you run the agent in.
#openai#gpt-5.6+6 more
2026-07-17
Read Article
Running Kimi K3 in Kimi Code: the Moderato plan unlocks 256K context, Allegretto the full 1M, and cache discipline — not knobs — controls your real cost.
#kimi k3#kimi code+5 more
2026-07-17
Read Article
Kimi K3's open weights are promised by July 27, not shipped — use the 10-day window to prep hosting for a 2.8T model and check the license before you commit.
#kimi k3#open-weight models+5 more
2026-07-17
Read Article
Kimi K3's open-weight bet meets Anthropic's closed Fable 5. Vendor benchmarks split 6-8, K3 lists at $3/$15 vs $10/$50, and weights are promised July 27.
#kimi k3#claude fable 5+5 more
2026-07-17
Read Article
Kimi K3 leads GPT-5.6 Sol on seven of fourteen vendor-reported benchmarks, but Sol's effort controls and ultra mode reframe agentic fit beyond raw scores.
#kimi k3#gpt-5.6 sol+5 more
2026-07-17
Read Article
Five open-weight moves cluster in one July window — K3, Inkling, M3 Pro, a Mistral MoE teaser and a scheduled DeepSeek V4 — narrowing the gap to a generation.
#open-weight models#kimi k3+5 more
2026-07-17
Read Article
Kimi K3 brings 2.8T parameters, 1M context, and near-frontier vendor benchmarks, with open weights due July 27, 2026. What the release means for AI buyers.
#kimi k3#moonshot ai+5 more
2026-07-17
Read Article
SpaceXAI open-sourced Grok Build's Rust harness under Apache 2.0 days after a privacy scandal. Why a public repo is not the same as a security audit.
#AI Development#Grok Build+5 more
2026-07-16
Read Article
Stay Ahead of the Curve

Marketing Insights Scrolled Straight to Your Inbox

Join 15,000+ marketers getting our weekly deep dives on SEO, AI trends, and growth strategies. No fluff, just actionable tactics.

View Our Services

Join a community of forward-thinking marketers. Unsubscribe at any time.