TPS
—
Average latency
—
Success rate
—
Pricing
Base Price
Input
—
Output
—
Pricing by Group
| Group | Ratio | Input | Output |
|---|
Model
Provider
Type
Groups
Endpoints
Capabilities / Supported modalities
Input
Output
Model ID
TPS
—
Sustained tokens per second
Average latency
—
Success rate
—
Per-group performance
Average latency, TTFT, TPS, and success rate
| Group | TPS | Average TTFT | Average latency | Success rate |
|---|
Latency trend (last 24h)
Average TTFT
Availability (last 24h)
Code samples
Authentication
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
Supported parameters
| Parameter | Type | Default / range | Description |
|---|
Rate limits
| Group | RPM | TPM | RPD |
|---|
RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply per token group.