Input budgets include the prompt template and an assumed amount of supplied text. The lower and upper amounts use the same input budget and different answer lengths. These are shared token scenarios for comparison; actual tokenization differs by model, language and content. Matching listed rates produce the same illustrative estimate, not identical token usage or quality.
Monthly savings = the listed or assumed monthly subscription price minus 20 times the per-prompt cost. The savings range reverses the cost range; percentages divide savings by the subscription price. Negative savings mean extra cost. Figures assume 20 independent runs with the same token budget, not an ongoing conversation. For tutoring and languages, these are 20 first responses, not complete learning sessions. Savings are only achievable if you can stop paying for the whole subscription and no longer need its other features. Annual payments may only be avoidable at renewal. Costs assume uncached text input, Standard processing and short-context OpenAI pricing. Batch, Flex, cache discounts, Fast mode and regional premiums are not included. OpenAI long-context requests can cost more; see the linked model pages. Extra reasoning tokens, system messages beyond the allowance, tool calls, generated files, images, audio, retries, follow-up conversation history, taxes and provider markups can increase the bill. Chat subscription fees and limits are separate. Tutoring and language estimates cover the first response, not the complete conversation.
DeepSeek peak hours are Monday–Friday, 01:00–04:00 and 06:00–10:00 UTC, excluding Chinese public holidays; all other hours, weekends and those holidays are off-peak. Both rates are shown. Use deepseek-flash for V4.1 Flash. Thinking is enabled by default, so additional billed reasoning can materially increase costs.