Back to providers



Cerebras
HealthyTelemetry updated 56m ago
Ultra-fast inference provider powered by the Cerebras Wafer Scale Engine, known for extremely high tokens/sec throughput. Offers OpenAI-compatible API with free tier access.
Supported Regions
us-east-1
Latency (TTFT)
Time to first token percentiles
No latency data available
Health History
Uptime over the last 7 days
7-Day Uptime100.00% — Excellent
24-Hour Uptime100.00% — Excellent
Current Status
Healthy
Last Checked
56m ago
Supported Models (3)
Models available through this provider. Select a model to view details.
K2 Think
k2-think
- Pricing (per 1M)
- In: $0.60
Out: $0.60 - Rate Limits
- 30 RPM
60K TPM - Regions
- us-east-1
Llama 4 Scout
llama-4-scout
- Pricing (per 1M)
- In: $0.60
Out: $0.60 - Rate Limits
- 30 RPM
60K TPM - Regions
- us-east-1
Llama 3.3 70B Instruct
llama-3-3-70b
- Pricing (per 1M)
- In: $0.60
Out: $0.60 - Rate Limits
- 30 RPM
60K TPM - Regions
- us-east-1