Fields
| Field | Type | Required | Description | Example |
|---|---|---|---|---|
ContextLength | int64 | ✅ | N/A | |
LatencyLast30m | *components.PercentileStats | ✅ | Latency percentiles in milliseconds over the last 30 minutes. Latency measures time to first token. Only visible when authenticated with an API key or cookie; returns null for unauthenticated requests. | { “p50”: 25.5, “p75”: 35.2, “p90”: 48.7, “p99”: 85.3 } |
MaxCompletionTokens | *int64 | ✅ | Maximum completion tokens for this endpoint. Input and output tokens share the context window, so the effective maximum output for a request is further limited by the context remaining after input tokens. | |
MaxPromptTokens | *int64 | ✅ | N/A | |
ModelID | string | ✅ | The unique identifier for the model (permaslug) | openai/gpt-4 |
ModelName | string | ✅ | N/A | |
Name | string | ✅ | N/A | |
NativeTools | map[string]components.NativeTools | ✅ | The server tools this endpoint accepts as the provider’s own built-in tool (engine: "native") instead of an OpenRouter engine, keyed by canonical openrouter:* name. Each value names the provider tool type the request is translated to. Where that tool runs (provider-side, or returned to the client as with Anthropic bash) is documented per tool. Empty when the provider has none. | { “openrouter:web_search”: { “type”: “web_search_20260209” } } |
PerfLast30mByWorkload | *components.PerfLast30mByWorkload | ➖ | Endpoint performance over the last 30 minutes, keyed by the kind of request served (e.g. text_generation, image_generation). Additive to the legacy singular latency and throughput fields; image and video generation report end-to-end latency. Only visible when authenticated with an API key or cookie. | |
Pricing | components.Pricing | ✅ | N/A | |
ProviderName | components.ProviderName | ✅ | N/A | OpenAI |
Quantization | *components.Quantization | ✅ | N/A | fp16 |
Status | *components.EndpointStatus | ➖ | N/A | 0 |
SupportedParameters | []components.Parameter | ✅ | N/A | |
SupportsImageReference | *bool | ➖ | Whether this TTS endpoint accepts an image_url reference describing the desired voice. Requests carrying an image reference are only routed to endpoints where this is true. | |
SupportsImplicitCaching | bool | ✅ | N/A | |
SupportsMultipleAudioReferences | *bool | ➖ | Whether this TTS endpoint accepts more than one input_audio reference clip per request. Requests carrying several clips are only routed to endpoints where this is true. | |
SupportsToolChoice | components.ToolChoiceSupport | ✅ | Per-variant tool_choice support. tool_choice in supported_parameters only says the parameter is accepted; these flags say which of its values passed testing. | { “auto”: true, “function”: true, “none”: true, “required”: true } |
SupportsVoiceCloning | *bool | ➖ | Whether this TTS endpoint accepts inline reference audio (input_references) for stateless voice cloning. Requests carrying reference audio are only routed to endpoints where this is true. | |
Tag | string | ✅ | N/A | |
ThroughputLast30m | *components.PercentileStats | ✅ | N/A | { “p50”: 25.5, “p75”: 35.2, “p90”: 48.7, “p99”: 85.3 } |
UptimeLast1d | *float64 | ✅ | Uptime percentage over the last day: the share of minutes in which at least 80% of provider attempts succeeded, counting only minutes with 10 or more attempts. Rate-limited and caller-caused failures are excluded. Null when no minute had enough traffic. | |
UptimeLast30m | *float64 | ✅ | N/A | |
UptimeLast5m | *float64 | ✅ | Uptime percentage over the last 5 minutes: the share of minutes in which at least 80% of provider attempts succeeded, counting only minutes with 10 or more attempts. Rate-limited and caller-caused failures are excluded. Null when no minute had enough traffic. |