Posts

AI Subscription Capacity Is Perishable

Why AI subscription limits are perishable, how purchased Codex resets differ from credits and grants, and how to use capacity without waste.

Posts

Prompt Caching: Cut AI Agent Token Costs

Learn how prompt caching changes LLM token costs, why coding agents resend context, what compaction can waste, and how to measure Codex, Claude Code, and OpenCode.

Models

Claude Sonnet 5 Guide

Claude Sonnet 5 pricing, 1M context, API migration changes, tokenizer cost implications, availability, and source-labeled launch evidence.

Models

GPT-5.6 Sol, Terra, and Luna Guide

Historical GPT-5.6 Sol, Terra, and Luna access, short- and long-context pricing, independent benchmark signals, and deployment controls.

Models

Claude Opus 4.8 Historical Guide

Claude Opus 4.8 historical pricing and benchmark context, retained for older integrations after Opus 5 became the premium Claude baseline.