Unbiased / unbiased/pareto

Pareto - access through LLMTR

Pareto is Unbiased's blended model: you send one request to a single model id and get one answer. It takes text and image input and returns text, with a 262,144-token context window and up to 131,072 output tokens. Images can be sent as base64 data URLs or https URLs. Tool calling works with tool_choice auto, including more than one call in a turn; send no other tool_choice value, and leave out the tools field when you do not want a tool used. response_format json_object is supported, json_schema is not. Repeated prompt prefixes are cached automatically and billed at the cached-input rate. Token counts for the same prompt can differ between calls, so read the cost from the usage of each response rather than estimating it from prompt length.

Technical specifications

Canonical IDunbiased/pareto
ProviderUnbiased
Context window262,144 tokens
OperationsCHAT_COMPLETIONS
Modalitiestext, image

Pricing

An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.

OperationMetricUnitPrice
CHAT_COMPLETIONSINPUT_TEXTPER_1M_TOKENS$2.50
CHAT_COMPLETIONSCACHE_READPER_1M_TOKENS$0.250000
CHAT_COMPLETIONSOUTPUT_TEXTPER_1M_TOKENS$7.50

Example usage

With existing OpenAI SDK flows, change only the base URL and model identifier.

curl https://llmtr.com/v1/chat/completions   -H "Authorization: Bearer llmtr-your_key"   -H "Content-Type: application/json"   -d '{"model":"unbiased/pareto","messages":[{"role":"user","content":"Hello"}]}'

Related models