* chore(models): fleet-wide provider catalog curation * fix(models): sync provider runtime catalogs and tests with curated manifests * test(models): align venice lifecycle assertions and copilot auth default with curated catalogs * test(models): align catalog lifecycle contract checks
4.6 KiB
summary, read_when, title
| summary | read_when | title | ||
|---|---|---|---|---|
| Use DeepInfra's unified API to access the most popular open source and frontier models in OpenClaw |
|
DeepInfra |
DeepInfra routes requests to popular open source and frontier models behind a single OpenAI-compatible endpoint and API key. Most OpenAI SDKs work against it by switching the base URL.
Install plugin
openclaw plugins install @openclaw/deepinfra-provider
openclaw gateway restart
Get an API key
- Sign in at deepinfra.com
- Go to Dashboard / Keys and generate a key, or use the auto-created one
CLI setup
openclaw onboard --deepinfra-api-key <key>
Or set the environment variable:
export DEEPINFRA_API_KEY="<your-deepinfra-api-key>" # pragma: allowlist secret
Config snippet
{
env: { DEEPINFRA_API_KEY: "<your-deepinfra-api-key>" }, // pragma: allowlist secret
agents: {
defaults: {
model: { primary: "deepinfra/deepseek-ai/DeepSeek-V4-Flash" },
},
},
}
Supported surfaces
Chat, image generation, and video generation refresh their model catalogs
live from https://api.deepinfra.com/v1/openai/models?sort_by=openclaw&filter=with_meta
once DEEPINFRA_API_KEY is configured. Live discovery expands the list of
selectable models; the default model per surface stays the static value
below. Other surfaces use static catalogs until they move onto the same
live catalog.
| Surface | Default model | OpenClaw config/tool |
|---|---|---|
| Chat / model provider | deepseek-ai/DeepSeek-V4-Flash (live catalog adds more chat models) |
agents.defaults.model |
| Image generation/editing | black-forest-labs/FLUX-1-schnell (live catalog adds more image-gen models) |
image_generate, agents.defaults.mediaModels.image |
| Media understanding | moonshotai/Kimi-K2.5 for images |
inbound image understanding |
| Speech-to-text | openai/whisper-large-v3-turbo |
inbound audio transcription |
| Text-to-speech | hexgrad/Kokoro-82M |
tts.provider: "deepinfra" |
| Video generation | Pixverse/Pixverse-T2V (live catalog adds more video-gen models) |
video_generate, agents.defaults.mediaModels.video |
| Memory embeddings | BAAI/bge-m3 |
memory.search.provider: "deepinfra" |
DeepInfra also exposes reranking, classification, object-detection, and other native model types. OpenClaw has no provider contract for those categories yet, so this plugin does not register them.
Available models
OpenClaw discovers DeepInfra models dynamically once a key is configured. Use
/models deepinfra or openclaw models list --provider deepinfra to see the
current list.
Any model on deepinfra.com works with the
deepinfra/ prefix:
deepinfra/deepseek-ai/DeepSeek-V4-Flash
deepinfra/deepseek-ai/DeepSeek-V4-Pro
deepinfra/zai-org/GLM-5.2
deepinfra/stepfun-ai/Step-3.7-Flash
deepinfra/moonshotai/Kimi-K2.7-Code
deepinfra/moonshotai/Kimi-K2.6
deepinfra/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B
deepinfra/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B
...and many more
Notes
- Model refs are
deepinfra/<provider>/<model>(for exampledeepinfra/Qwen/Qwen3-Max). - Default chat model:
deepinfra/deepseek-ai/DeepSeek-V4-Flash - Base URL:
https://api.deepinfra.com/v1/openai - Video generation uses the OpenAI-compatible async endpoint
https://api.deepinfra.com/v1/openai/videos(submit, then poll). A configuredbaseUrlis honored.openclaw doctor --fixmigrates legacynativeBaseUrlor/v1/inferencevalues onapi.deepinfra.comtobaseUrlautomatically; custom native endpoints are retired with a doctor notice and need a manually configured OpenAI-compatiblebaseUrl. Video generation fails with an actionable error (before sending any request) whilebaseUrlstill targets the retired/v1/inferencesurface.