Model catalog

Use the catalog to decide which model ID to send in inference requests. The public model browser and catalog endpoint show the same customer-facing data.

Endpoints

GET/v1/modelsOpenAI-compatible model list.
GET/v1/models/catalogFull catalog with providers, tags, pricing, context windows, and release dates.
GET/v1/analytics/latencyPublic latency aggregates by model and provider.

Providers

Mistral AI

Chat, embedding, transcription, and TTS models.

Scaleway

EU GPU inference for selected open models.

OVHcloud

EU cloud inference for selected models, including image generation.

Inceptron

OpenAI-compatible inference from Sweden for selected large models.

Browse models visually at /models.