Meta / meta/llama-3.3-70b-instruct-12k
Llama 3.3 70B Instruct (12K) - access through LLMTR
Llama 3.3 70B Instruct is Meta's multilingual open-weight instruction model. It is a strong general-purpose option for chat, summarization, content drafting, coding help, and function-calling assistant flows. Both the context window and the maximum single-response output are capped at 12,288 tokens. Supported languages: English, German, French, Italian, Portuguese, Hindi, Spanish, and Thai.
Technical specifications
| Canonical ID | meta/llama-3.3-70b-instruct-12k |
|---|---|
| Provider | Meta |
| Context window | 12,288 tokens |
| Operations | CHAT_COMPLETIONS |
| Modalities | text |
Pricing
An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.
| Operation | Metric | Unit | Price |
|---|---|---|---|
| CHAT_COMPLETIONS | INPUT_TEXT | PER_1M_TOKENS | $0.135000 |
| CHAT_COMPLETIONS | OUTPUT_TEXT | PER_1M_TOKENS | $0.400000 |
Example usage
With existing OpenAI SDK flows, change only the base URL and model identifier.
curl https://llmtr.com/v1/chat/completions \
-H "Authorization: Bearer llmtr-your_key" \
-H "Content-Type: application/json" \
-d '{"model":"meta/llama-3.3-70b-instruct-12k","messages":[{"role":"user","content":"Hello"}]}'
Related models
- Muse Spark 1.2 meta/muse-spark-1.2
- Muse Glimmer 30B meta/muse-glimmer-30b
- Muse Spark 1.1 meta/muse-spark-1.1
- Muse Spark 1.2 Contributor meta/muse-spark-1.2-contributor
- Gemma 4 llmtr/gemma-4
- Qwen 3.6 35B-A3B llmtr/qwen3-6-35b
- Trendyol Asure 12B llmtr/trendyol-asure-12b
- Magibu 11B v8 llmtr/magibu-11b-v8