Skip to content

Hermes 4 405B

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research.

Per 1M tokens
$1.05 · $3.15
Context
131Ktokens
Supports
  • Thinks before it answers
opencode.json
{
  "provider": {
    "deference": {
      "npm": "@ai-sdk/openai-compatible",
      "options": { "baseURL": "https://deference.si/v1", "apiKey": "{env:DEFERENCE_API_KEY}" },
      "models": { "nousresearch/hermes-4-405b": { "limit": { "context": 131072, "output": 117964 } } }
    }
  }
}
Needs an API keyGet one

Pricing and limits

  • Max output117,964tokens

OpenRouter's price plus OpenRouter's 5% platform fee. Deference adds no markup.