AI Model Context & Character Limits

Every model accepts a prompt only up to a certain size. Send more than its context window and it either truncates your prompt or refuses. This table shows the real limits and per-token prices for the models people ask about most, taken directly from the 191 models in the TPEE app (last updated 2026-08-25 09:30 +10:00).

By Jacques Botte, founder of Toptronic®. Last updated 12 September 2026.

Context windows and prices (headline models)

ModelContext windowInput $/1MOutput $/1MReleased
Kimi K31M$3$1507-16-2026
Claude Fable 51M$10$5006-09-2026
GPT-5400K$1.25$1008-07-2025
Gemini 3.5 Flash1M$1.5$905-19-2026
DeepSeek V4 Pro1M$0.435$0.8704-24-2026
Qwen 3.7 Max1M$2.5$7.506-01-2026
Mistral Large 3256K$0$005-20-2026
Grok 4.6500K$2$608-12-2026
GLM-5202K$1$3.202-12-2026
Llama 4 Scout10M$0.08$0.304-05-2025

Prices and windows change often — this table is a snapshot. The full, up-to-date list of all 191 models across 20 providers ships inside the TPEE app, where the API Character Limits panel checks your prompt against the exact limit of the model you've chosen before you send it.

Why the limit matters

A prompt that's too long for the model gets cut off mid-thought — the model never sees the end of your instructions and answers as if they weren't there. A prompt that's just under the limit, meanwhile, may leave no room for the model's own answer. TPEE's character counters and per-model limits let you size the prompt to the model you're actually paying for, so you don't discover the problem after you've spent the tokens.

Get TPEE or read the prompt engineering guide.