AI & LLM
LLM tokens, API pricing and context windows — FastMinify guides for OpenAI, Claude and Gemini, 100% browser-local.

LLM API Cost: Batch API, Prompt Caching and Monthly Projection
Lower your OpenAI/Anthropic/Gemini bill: estimate input/output, model batch and caching in your calculations — verified rates, 100% local.
21.08.2026
10 min read
llm
pricing
batch-api
+4

LLM Context Window: Size Your Prompts for GPT, Claude and Gemini
RAG, multi-turn agents, system prompts: calculate context-window usage and remaining headroom before calling the API.
19.08.2026
7 min read
llm
context-window
prompt
+4

When TOON Beats JSON for LLM Prompts (and When It Doesn't)
Honest guide: uniform tabular data, convert → count → price → context workflow; not a JSON/YAML replacement manifesto.
16.08.2026
8 min read
toon
JSON
llm
+4

Count LLM Tokens and Estimate API Cost Locally (OpenAI, Claude, Gemini)
Before calling an LLM API: count tokens, estimate USD cost and check context-window fit — 100% in the browser without sending your prompt.
14.08.2026
7 min read
llm
tokens
pricing
+5