GreenPT / greenpt/deepseek-v4-flash-0731

DeepSeek V4 Flash - access through LLMTR

DeepSeek V4 Flash pairs a 1,000,000-token context window with one of the lowest prices on this provider. Tool/function calling was verified over a full two-turn loop. Repeated prompt prefixes are served from a cache and those input tokens bill at roughly a quarter of the normal input rate; in measurement the cache took a few calls to warm, so do not conclude from a single repeat that there is none. It does not accept image input; a request containing an image is refused by LLMTR without being charged. The model reasons by default and those tokens bill as output.

Technical specifications

Canonical IDgreenpt/deepseek-v4-flash-0731
ProviderGreenPT
Context window1,000,000 tokens
OperationsCHAT_COMPLETIONS
Modalitiestext

Pricing

An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.

OperationMetricUnitPrice
CHAT_COMPLETIONSINPUT_TEXTPER_1M_TOKENS$0.170843
CHAT_COMPLETIONSCACHE_READPER_1M_TOKENS$0.048812
CHAT_COMPLETIONSOUTPUT_TEXTPER_1M_TOKENS$0.427109

Example usage

With existing OpenAI SDK flows, change only the base URL and model identifier.

curl https://llmtr.com/v1/chat/completions   -H "Authorization: Bearer llmtr-your_key"   -H "Content-Type: application/json"   -d '{"model":"greenpt/deepseek-v4-flash-0731","messages":[{"role":"user","content":"Hello"}]}'

Related models