Meta / meta/llama-3.3-70b-instruct-12k
Llama 3.3 70B Instruct (12K) - access through LLMTR
Llama 3.3 70B Instruct is Meta's multilingual open-weight instruction model. It is a strong general-purpose option for chat, summarization, content drafting, coding help, and function-calling assistant flows. Both the context window and the maximum single-response output are capped at 12,288 tokens. Supported languages: English, German, French, Italian, Portuguese, Hindi, Spanish, and Thai.
Technical specifications
| Canonical ID | meta/llama-3.3-70b-instruct-12k |
|---|---|
| Provider | Meta |
| Context window | 12,288 tokens |
| Operations | CHAT_COMPLETIONS |
| Modalities | text |
Pricing
An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.
| Operation | Metric | Unit | Price |
|---|---|---|---|
| CHAT_COMPLETIONS | INPUT_TEXT | PER_1M_TOKENS | $0.135000 |
| CHAT_COMPLETIONS | OUTPUT_TEXT | PER_1M_TOKENS | $0.400000 |
Example usage
With existing OpenAI SDK flows, change only the base URL and model identifier.
curl https://llmtr.com/v1/chat/completions -H "Authorization: Bearer llmtr-your_key" -H "Content-Type: application/json" -d '{"model":"meta/llama-3.3-70b-instruct-12k","messages":[{"role":"user","content":"Hello"}]}'
Related models
- Muse Spark 1.3 meta/muse-spark-1.3
- Muse Glimmer 30B meta/muse-glimmer-30b
- Muse Image meta/muse-image-1.0
- Muse Spark 1.2 meta/muse-spark-1.2
- Muse Spark 1.1 meta/muse-spark-1.1
- Muse Spark 1.2 Contributor meta/muse-spark-1.2-contributor
- Muse Spark 1.3 Contributor meta/muse-spark-1.3-contributor
- Llama 3.1 8B Instruct meta/llama-3.1-8b-instruct-16k