Get started
Choose a model
Find a model id, check its price and context, and send it to the right endpoint.
A model id has the form author/name, for example anthropic/claude-sonnet-5.5. Send it as model on every endpoint.
Browse the catalog
Models lists every model with its context length, price per million tokens and capabilities. Open a model to copy its id and a ready-to-run example.
Pick a maker in the maker bar, choose a capability in the Capabilities menu, or link straight to a filter:
| Capability | Link | Shows |
|---|---|---|
| All | /models | Every listed model |
| Reasoning | /models?type=reasoning | Models that think before they answer |
| Tools | /models?type=tools | Models that call functions |
| Vision | /models?type=vision | Models that read images |
| Images | /models?type=image | Models that create images |
| Embeddings | /models?type=embeddings | Models that turn text into vectors |
Add q= to search and provider= to filter by maker: /models?type=tools&provider=anthropic.
Match the model to the endpoint
| Model type | Endpoint |
|---|---|
| Text, vision, reasoning, tools | Chat completions, Messages or Responses |
| Embeddings | Embeddings |
| Image generation | Images |
Coding tools pick the endpoint themselves: Claude Code uses Messages, Codex uses Responses, and most other tools use Chat completions.
Read the price
Prices are in dollars per million tokens, with separate prices for input, output and cached input. Some models also charge per request or per image. A price is OpenRouter's price plus OpenRouter's 5% crypto platform fee, with no Deference markup. See Pricing.
Free credit pays for open-weight text models whose output costs at most $2 per million tokens. Every other model needs credit.
List models from the API
GET /v1/models needs no key and returns each model's id, context window, prices and supported parameters. See Models.
curl https://deference.si/v1/modelsUse aliases and suffixes
- Ids that start with
~follow the newest model of a family. - A trailing
[1m]is accepted so Claude Code's context setting works. Deference ignores it when it looks up the model and its price, and forwards the id as you wrote it.
If a model is missing
A model that is not in the catalog returns 404 model_not_found. Providers retire models, so list models again and pick a current id.