The PromptCostLab journal

Useful notes for building with AI.

Model guides, cost comparisons, and plain-English explanations you can put to work immediately.

Built-in guides are active. Connect Sanity to publish from your own editorial workspace.

Model guides · 8 min

Claude Fable 5 token limits, pricing, and context fit

A plain-English guide to planning prompts for Anthropic's flagship family, including thinking-token billing and cache pricing.

Read guide →
Token basics · 6 min

Why the same prompt has different token counts

Tokenizers explained without the textbook: vocabulary, chunks, and hidden request overhead.

Read guide →
Registry · 5 min

The 2026 model registry: what changed this month

A living update log for major model families, context windows, and pricing snapshots.

Read guide →
Token basics · 6 min

A beginner's guide to context windows

The difference between tokens, context, output limits, and what actually gets truncated.

Read guide →
Token basics · 9 min

Tokens to Words: How Many Words Is 100, 1,000, or 128K Tokens?

Conversion tables for every common token count, the 0.75 rule behind them, and why code and other languages break it.

Read guide →
Guides · 10 min

Context Windows Explained: Every Major Model Compared (2026)

What a context window actually contains, how 200K to 1M tokens translate into books and codebases, and the current window for every major model.

Read guide →
Guides · 8 min

Input, Output, and Cached Tokens: What You're Actually Billed For

How input tokens are calculated, what cached input really costs, and the billing categories that surprise first-time API users.

Read guide →
Model guides · 9 min

Claude Code Cost: Subscription vs API Billing, Calculated

What Claude Code really costs on Pro, Max, and Team plans versus pay-per-token API mode — with the numbers to decide which fits your usage.

Read guide →
Token basics · 6 min

Tokens to Characters: How Many Characters Is a Token?

The character-to-token ratios for English, code, and other languages, with conversion tables and when to trust them.

Read guide →
Model guides · 7 min

Grok API Cost Explained: Grok 4.6 vs 4.5 Pricing

Current xAI token pricing including cached input, the 200K long-context doubling, and what a realistic Grok workload costs per month.

Read guide →
Token basics · 8 min

The AI Token Glossary: Every Token Term Defined

Plain-English definitions for tokens, tokenizers, context windows, caching, and every billing term on an LLM invoice.

Read guide →
Token basics · 6 min

Words to Tokens: Estimate Prompt Size Before You Paste

How to convert a word count into a token budget — the 1.3 rule, per-content corrections, and a converter that does it for you.

Read guide →
Pricing · 7 min

How Much Does 1 Million Tokens Cost? (2026 LLM Pricing)

The price of a million input and output tokens on every major model — plus what a million tokens actually holds.

Read guide →
Comparisons · 9 min

Kimi K3 vs Claude Sonnet: which prompt costs less?

A practical comparison for long-context chat, RAG, and production assistants.

Read guide →
Optimization · 8 min

How to cut your AI bill without making prompts worse

Five practical ways to reduce tokens while keeping the instructions that matter.

Read guide →