skip to content- 2026-09-27
|
AI Agents Crossed the Line: What the Incidents Show
Hugging Face, Australia's Medicare portal, Gemini and Claude: confirmed scope, new SwarmTraces evidence, and practical lessons for anyone deploying AI agents.
- 2026-09-27
|
Jev: Practical Use Cases and a Builder’s Guide
Build with TypeSafe Jev: useful projects, OpenRouter access, agent skills, an interactive routing example, and an evaluation plan that separates promise from proof.
- 2026-09-27
|
Open Weights vs Open Source: What MiMo V2.6 Actually Gives You
A plain-English guide to weights, model code, RL environments, licenses, and reproducibility, using Xiaomi MiMo V2.6 as a dated example.
- 2026-09-27
|
How to Read Xiaomi MiMo's RL Training Dashboard
A plain-English guide to Xiaomi MiMo V2.6's public RL dashboard: rollouts, rewards, batches, benchmarks, open weights, and compute limits.
- 2026-09-14
|
DeepSeek V4.1 Flash: A Strong API Value Candidate
DeepSeek V4.1 Flash adds native vision and MIT weights. Compare direct API pricing, legacy aliases, independent evidence, and temporary coding-plan offers.
- 2026-09-14
|
GPT-6 Astra: Start Planning Evaluations on Low
Evaluate GPT-6 Astra Low for planning and orchestration, with benchmark regressions, API prices, subscription limits, and accepted-result costs kept distinct.
- 2026-08-01
|
DeepSeek V4 Flash 0731: The API Value Frontier
DeepSeek V4 Flash 0731 pairs a 1M context window and MIT weights with a $0.14/$0.28 direct API, but high token use and hallucinations still require evaluation.
- 2026-08-01
|
GPT-5.6 Luna’s July Price Cut: Historical Analysis
Why GPT-5.6 Luna became the July OpenAI value candidate for bounded Codex work and high-volume API tasks, with subscription and API economics kept separate.
- 2026-07-26
|
Claude vs OpenAI Data Retention: What Gets Kept
Claude and OpenAI retention compared by consumer, API, Covered Model, ZDR, safety-review, legal-hold, and local deployment paths.
- 2026-07-26
|
AI Subscription Capacity Is Perishable
Why AI subscription limits are perishable, how purchased Codex resets differ from credits and grants, and how to use capacity without waste.
- 2026-07-26
|
Prompt Caching: Cut AI Agent Token Costs
Learn how prompt caching changes LLM token costs, why coding agents resend context, what compaction can waste, and how to measure Codex, Claude Code, and OpenCode.
- 2026-07-25
|
How OpenAI's Cyber Eval Breached Hugging Face
The July preliminary, source-labeled analysis of the Hugging Face breach, with September follow-up evidence and five cyber-agent containment controls.
- 2026-07-02
|
Fable 5 Returns as Claude Sonnet 5 Becomes Default
Fable 5 is back with tighter safeguards while Sonnet 5 becomes Claude's default. Here is the practical routing and cost decision.
- 2026-07-02
|
How to Read AI Benchmarks Without Getting Fooled
Learn how to read AI benchmarks, spot saturation and contamination, compare coding and chat leaderboards, and test models on your own work.
- 2026-06-28
|
GPT-5.6 Sol Preview: Access and Agent Risks
GPT-5.6 Sol is a restricted API and Codex preview shaped by a government request. Check access, agent risks, pricing, and safeguards before routing work.
- 2026-06-28
|
About
Why this site exists and how it is run.
- 2026-06-27
|
Frontier Model Access: Fable, Mythos, GPT-5.6
Fable 5 is restored globally, while Mythos 5 and GPT-5.6 remain approval-sensitive. These are different access mechanisms, not one licensing regime.
- 2026-06-27
|
Codex Banked Resets and the GPT-5.6 Preview
A dated Codex chronology separating scheduled limits, banked resets, purchased full resets, credits, incident recovery, and broad resets through August 20.
- 2026-06-11
|
Claude Fable 5 Restored: Cost and Guardrails
Fable 5 is generally available but credit-controlled. Compare its $10/$50 price, safeguards, retention, and Opus 5 alternative.
- 2026-05-21
|
AI Coding Subscription Limits: May 2026
May 2026 coding-limit snapshot. Current discovery routes Kimi K3 as the newest flagship and K2.7 Code as the cheaper coding API.