StepFun / stepfun/step-3.7-flash
Step 3.7 Flash - access through LLMTR
Step 3.7 Flash is a multimodal reasoning model for text, image, and video input through Chat Completions. It is suited to coding help, multi-step agent flows, document or screenshot interpretation, and long-context analysis; reasoning_effort low, medium, and high tune the latency/depth balance.
Technical specifications
| Canonical ID | stepfun/step-3.7-flash |
|---|---|
| Provider | StepFun |
| Context window | 256,000 tokens |
| Operations | CHAT_COMPLETIONS |
| Modalities | text, image, video |
Pricing
An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.
| Operation | Metric | Unit | Price |
|---|---|---|---|
| CHAT_COMPLETIONS | INPUT_TEXT | PER_1M_TOKENS | $0.200000 |
| CHAT_COMPLETIONS | CACHE_READ | PER_1M_TOKENS | $0.040000 |
| CHAT_COMPLETIONS | OUTPUT_TEXT | PER_1M_TOKENS | $1.15 |
Example usage
With existing OpenAI SDK flows, change only the base URL and model identifier.
curl https://llmtr.com/v1/chat/completions -H "Authorization: Bearer llmtr-your_key" -H "Content-Type: application/json" -d '{"model":"stepfun/step-3.7-flash","messages":[{"role":"user","content":"Hello"}]}'
Guides about this model
- Building a video analysis flow with video input and reasoning_effort on Step 3.7 Flash - Step 3.7 Flash is one of the few models combining image and video input with thinking-level control. This article covers which reasoning_effort level to prefer, and when, in a video analysis task.
Related models
- Step 3.5 Flash stepfun/step-3.5-flash
- Step 3.5 Flash 2603 stepfun/step-3.5-flash-2603
- Gemma 4 llmtr/gemma-4
- Qwen 3.6 35B-A3B llmtr/qwen3-6-35b
- Trendyol Asure 12B llmtr/trendyol-asure-12b
- Magibu 11B v8 llmtr/magibu-11b-v8
- Muse Glimmer 30B (Turkey) llmtr/muse-glimmer-30b-tr
- EmbeddingGemma 300M llmtr/embeddinggemma-300m