Skip to main content
Don’t want to host it yourself? DeepInfra can run Hermes for you — see Hosted Agents: Hermes-Agent.
Hermes Agent is a self-hosted, self-improving autonomous agent by Nous Research. DeepInfra is a built-in Hermes provider: set provider: deepinfra, export your key as DEEPINFRA_API_KEY, and pick any LLM from our catalog. Hermes reads the model list, context lengths and pricing live from the DeepInfra catalog, so there is no base URL or context length to configure.

Configure ~/.hermes/config.yaml

Get your API key from the Dashboard and export it, or put the same line in ~/.hermes/.env:
  • The DeepInfra model id passes through verbatim in default — no reformatting needed.
  • Reasoning is controlled through DeepInfra’s reasoning_effort field, so agent.reasoning_effort, /reasoning <level> and --reasoning work in both directions: an effort turns thinking on for models that default off (DeepSeek-V4.x), and /reasoning none turns it off for models that default on.
  • Hermes needs roughly 64k of context for agent functionality. Check a model’s window and pricing via /v1/openai/models?filter=with_meta&sort_by=hermes.
  • On a Hermes release older than the built-in provider, use provider: custom with base_url: https://api.deepinfra.com/v1/openai, api_key: ${DEEPINFRA_API_KEY} and an explicit context_length.

Run it

To try a model without editing the YAML, pass the provider and model on the command line:

Learn more

AI Providers

Hermes’ provider overview, including the DeepInfra section.

Configuration

The full config.yaml reference.

Chat Completions

DeepInfra’s OpenAI-compatible API.

NVIDIA OpenShell

Run Hermes in a sandbox whose only allowed destination is DeepInfra.
Tested with Hermes Agent v0.21.5 (2026.9.24).