Use a measured count
Count a representative prompt first, then add expected output tokens. Plain word counts are not a reliable substitute for model tokenization.
Open the token counterCheck whether a prompt and planned output fit inside a model's tracked context window. Enter token counts, choose a model, and see the remaining capacity before you send the request.
These limits come from the current model records. Availability, API route, output caps, and request formatting can still affect a real call.
| Provider | Model | Context window | Checked |
|---|---|---|---|
| OpenAI | GPT-4o mini | 128,000 tokens | 2026-08-02 |
| OpenAI | GPT-4o | 128,000 tokens | 2026-08-02 |
| OpenAI | GPT-5.6 Sol | 1,050,000 tokens | 2026-08-02 |
| OpenAI | GPT-5.6 Terra | 1,050,000 tokens | 2026-08-02 |
| OpenAI | GPT-5.6 Luna | 1,050,000 tokens | 2026-08-02 |
| OpenAI | gpt-5.4-mini | 400,000 tokens | 2026-08-02 |
| OpenAI | gpt-5.4-nano | 400,000 tokens | 2026-08-02 |
| Anthropic | Claude Sonnet 5 | 1,000,000 tokens | 2026-08-02 |
| Anthropic | Claude Opus 4.8 | 1,000,000 tokens | 2026-08-02 |
| Anthropic | Claude Sonnet 4.6 | 1,000,000 tokens | 2026-08-02 |
| Anthropic | Claude Haiku 4.5 | 200,000 tokens | 2026-08-02 |
| Gemini 2.5 Pro | 1,048,576 tokens | 2026-08-02 | |
| Gemini 2.5 Flash | 1,048,576 tokens | 2026-08-02 | |
| Gemini 2.5 Flash-Lite | 1,048,576 tokens | 2026-08-02 | |
| DeepSeek | DeepSeek V4 Pro | 1,000,000 tokens | 2026-08-02 |
| DeepSeek | DeepSeek V4 Flash non-thinking | 1,000,000 tokens | 2026-08-02 |
| Anthropic | Claude Fable 5 | 1,000,000 tokens | 2026-08-02 |
| Anthropic | Claude Mythos 5 | 1,000,000 tokens | 2026-08-02 |
| Gemini 3.5 Flash | 1,048,576 tokens | 2026-08-02 | |
| Gemini 3.5 Live Translate Preview | 131,072 tokens | 2026-08-02 | |
| Gemini 3 Flash Preview | 1,048,576 tokens | 2026-08-02 | |
| Gemini 3.1 Flash-Lite | 1,048,576 tokens | 2026-08-02 | |
| Gemini 3.1 Pro Preview | 1,048,576 tokens | 2026-08-02 | |
| Groq | GPT-OSS 20B on Groq | 131,072 tokens | 2026-07-15 |
| Groq | GPT-OSS 120B on Groq | 131,072 tokens | 2026-07-15 |
| Groq | Llama 4 Scout on Groq | 131,072 tokens | 2026-07-15 |
| Groq | Qwen3-32B on Groq | 131,072 tokens | 2026-07-15 |
| Cohere | Command R 08-2024 | 128,000 tokens | 2026-07-17 |
| Cohere | Command R+ 08-2024 | 128,000 tokens | 2026-07-17 |
| Cohere | Command A | 256,000 tokens | 2026-07-17 |
| Mistral | Mistral Medium 3.5 | 256,000 tokens | 2026-07-17 |
| Moonshot AI | Kimi K3 | 1,048,576 tokens | 2026-07-17 |
| Moonshot AI | Kimi K2.7 Code | 262,144 tokens | 2026-07-17 |
| Moonshot AI | Kimi K2.7 Code High-Speed | 262,144 tokens | 2026-07-17 |
| Moonshot AI | Kimi K2.6 | 262,144 tokens | 2026-07-17 |
| xAI | Grok 4.5 | 500,000 tokens | 2026-07-17 |
| Gemini 3.6 Flash | 1,048,576 tokens | 2026-08-02 | |
| Gemini 3.5 Flash-Lite | 1,048,576 tokens | 2026-08-02 | |
| Anthropic | Claude Opus 5 | 1,000,000 tokens | 2026-08-02 |
Count a representative prompt first, then add expected output tokens. Plain word counts are not a reliable substitute for model tokenization.
Open the token counterA prompt can fit several models but have very different input, output, cache, and long-context costs.
Estimate API costStackLens records the checked model limit and source date. Open the model page for the source note and documented caveats.
Find a modelIt compares the token count you enter with a model's tracked context-window limit and shows the remaining capacity. It does not tokenize the text itself.
A model request usually shares its context budget across input and generated output, but provider behavior can vary. Treat the result as a planning check and verify the selected model's official documentation.
Context limits differ by model and sometimes by API route or product tier. Use the model's source-tracked limit here, then check the provider documentation before sending a large request.