Back to providers




Groq
HealthyTelemetry updated 55m ago
Ultra-fast LPU (Language Processing Unit) inference provider offering extremely low latency. Supports streaming, function calling, and audio transcription via Whisper models. Per-model rate limits apply.
Supported Regions
us-east-1eu-west-1
Latency (TTFT)
Time to first token percentiles
No latency data available
Health History
Uptime over the last 7 days
7-Day Uptime100.00% — Excellent
24-Hour Uptime100.00% — Excellent
Current Status
Healthy
Last Checked
55m ago
Supported Models (4)
Models available through this provider. Select a model to view details.
Llama 4 Scout
llama-4-scout
- Pricing (per 1M)
- In: $0.11
Out: $0.34 - Rate Limits
- 30 RPM
100K TPM - Regions
- us-east-1eu-west-1
Qwen3 32B
qwen3-32b
- Pricing (per 1M)
- In: $0.29
Out: $0.39 - Rate Limits
- 30 RPM
100K TPM - Regions
- us-east-1
DeepSeek R1
deepseek-r1
- Pricing (per 1M)
- In: $0.75
Out: $0.99 - Rate Limits
- 30 RPM
100K TPM - Regions
- us-east-1
Llama 3.3 70B Instruct
llama-3-3-70b
- Pricing (per 1M)
- In: $0.59
Out: $0.79 - Rate Limits
- 30 RPM
100K TPM - Regions
- us-east-1eu-west-1