MiniMax / minimax/minimax-m3.1-flash-preview

MiniMax M3.1 Flash Preview - access through LLMTR

MiniMax M3.1 Flash Preview is MiniMax's closed-beta model for coding and agent workflows; it is not hosted in Turkey, and its behaviour may change during the beta. It provides a 1,000,000-token context, returns at most 524,288 tokens per response and accepts text and images. Tool calling and JSON object mode work; reasoning depth is set with reasoning_effort from low to max. It reasons before every answer and the reasoning counts toward the output limit, so do not set max_tokens low. The model is free and is not charged to your balance, but using it requires an account that has been topped up at least once (minimum $5). The free period ends on 6 October 2026 at 23:59 (TRT); after that the metered minimax/minimax-m3 is available.

Technical specifications

Canonical IDminimax/minimax-m3.1-flash-preview
ProviderMiniMax
Context window1,000,000 tokens
OperationsCHAT_COMPLETIONS
Modalitiestext, image

Pricing

An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.

OperationMetricUnitPrice
CHAT_COMPLETIONSINPUT_TEXTPER_1M_TOKENSNot available
CHAT_COMPLETIONSCACHE_READPER_1M_TOKENSNot available
CHAT_COMPLETIONSOUTPUT_TEXTPER_1M_TOKENSNot available

Example usage

With existing OpenAI SDK flows, change only the base URL and model identifier.

curl https://llmtr.com/v1/chat/completions   -H "Authorization: Bearer llmtr-your_key"   -H "Content-Type: application/json"   -d '{"model":"minimax/minimax-m3.1-flash-preview","messages":[{"role":"user","content":"Hello"}]}'

Guides about this model

Related models