Unbiased / unbiased/pareto
Pareto - access through LLMTR
Pareto is Unbiased's blended model: you send one request to a single model id and get one answer. It takes text and image input and returns text, with a 262,144-token context window and up to 131,072 output tokens. Images can be sent as base64 data URLs or https URLs. Tool calling works with tool_choice auto, including more than one call in a turn; send no other tool_choice value, and leave out the tools field when you do not want a tool used. response_format json_object is supported, json_schema is not. Repeated prompt prefixes are cached automatically and billed at the cached-input rate. Token counts for the same prompt can differ between calls, so read the cost from the usage of each response rather than estimating it from prompt length.
Technical specifications
| Canonical ID | unbiased/pareto |
|---|---|
| Provider | Unbiased |
| Context window | 262,144 tokens |
| Operations | CHAT_COMPLETIONS |
| Modalities | text, image |
Pricing
An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.
| Operation | Metric | Unit | Price |
|---|---|---|---|
| CHAT_COMPLETIONS | INPUT_TEXT | PER_1M_TOKENS | $2.50 |
| CHAT_COMPLETIONS | CACHE_READ | PER_1M_TOKENS | $0.250000 |
| CHAT_COMPLETIONS | OUTPUT_TEXT | PER_1M_TOKENS | $7.50 |
Example usage
With existing OpenAI SDK flows, change only the base URL and model identifier.
curl https://llmtr.com/v1/chat/completions -H "Authorization: Bearer llmtr-your_key" -H "Content-Type: application/json" -d '{"model":"unbiased/pareto","messages":[{"role":"user","content":"Hello"}]}'
Related models
- Gemma 4 llmtr/gemma-4
- Qwen 3.6 35B-A3B llmtr/qwen3-6-35b
- Trendyol Asure 12B llmtr/trendyol-asure-12b
- Magibu 11B v8 llmtr/magibu-11b-v8
- Muse Glimmer 30B (Turkey) llmtr/muse-glimmer-30b-tr
- Leyla llmtr/leyla
- EmbeddingGemma 300M llmtr/embeddinggemma-300m
- Qwen 3.5 4B llmtr/qwen3-5-4b