Is Cerebras down right now?
Authenticated API inference - 2 models monitored · How we classify outages
Weekly AI pricing & uptime digest
Price drops, new model releases, and incident summaries - every Monday. Free.
Cerebras is currently down - API inference failing (5+ consecutive checks) - 374ms HTTP response. Last checked . 90-day uptime: 99.9%. Cerebras API: 0/2 models up - check inference section below.
llama3.1-8b API Outage
Stay informed
HTTP uptime (90d)
99.9%
2 incidents (90d)
HTTP response now
374ms
HTTP p50 (7d)
587ms
median ping response
HTTP p95 (7d)
1149ms
tail ping response
API Inference Monitoring
Live · every 5 minBest TTFT (p50)
—
time to first token
Best throughput
—
output tokens/sec (24h avg)
Min success rate
0%
worst model (24h)
P50 = typical speed. P95 = worst case 95% of the time. Measured by Tickerr's independent inference checks. Requires ≥10 checks to display.
TTFT over 24 hours
ⓘ Authenticated streaming API calls via native fetch. TTFT = milliseconds from request start to first streamed token chunk. Throughput = output tokens ÷ generation time. Checks run from Vercel us-east-1. Independent of the provider's official status page.
Is Cerebras slow right now?
Cerebras is currently experiencing an outage or significant issues. Check the status section above for details.Last checked 3m ago.
HTTP endpoint response: 374ms (7-day p50: 587ms). This measures basic reachability, not model inference speed.
TTFT (time-to-first-token) is measured via authenticated streaming API calls every 5 minutes. Current values are compared against 7-day and 30-day rolling baselines. A minimum of 5 successful checks is required before classification. Methodology
Cerebras speed FAQ
Is Cerebras slow right now?
Cerebras is currently experiencing an outage or significant issues. Check the status section above for details.
Why is Cerebras so slow today?
Cerebras is currently experiencing an outage, not just slowness. Check the status section above for the latest incident details and expected resolution.
Is Cerebras down or just responding slowly?
Cerebras is currently down — not just slow. An active outage or major incident has been detected. Check the status section above for details.
How does Tickerr determine whether Cerebras is slow?
Tickerr sends authenticated API requests to Cerebras every 5 minutes and measures time-to-first-token (TTFT). Current TTFT is compared against a rolling 7-day and 30-day baseline. A model is classified as slow when its current p50 TTFT exceeds 2x its baseline, and severely slow at 3.5x. At least 5 successful checks are required before any classification is made.
What is normal response time for Cerebras?
Tickerr is still collecting baseline data for Cerebras. Normal response times will be available after sufficient monitoring history has accumulated.
Can one Cerebras model be slow while others are working normally?
Yes. Cerebras runs multiple models (e.g. llama3.1-8b, qwen-3-235b-a22b-instruct-2507), and each can have independent performance characteristics. Tickerr monitors each model separately. A slowdown on one model doesn't necessarily affect others.
Experimental capability checks
PUBLIC PILOTTickerr sends deterministic test prompts to Cerebras every hour and evaluates the output using code-based pass/fail checks. These results are informational onlyduring the pilot period and do not affect the tool's status verdict.
How Tickerr tests this
Every hour, Tickerr sends three deterministic test prompts to each monitored model via the provider's API. Each response is evaluated with a code-based pass/fail check — no subjective judgement.
- Instruction following: asks the model to reply with an exact word. Graded by exact string match.
- JSON schema: asks for a JSON object matching a fixed schema. Graded by parse + field validation.
- Tool call: asks the model to call a function with specific arguments. Graded by parsing the tool-call response.
Infrastructure errors (timeouts, rate limits, auth failures) are tracked separately and excluded from capability pass rates. Models that don't support tool calls are marked “not supported” rather than counted as failures. Full methodology
HTTP endpoint response time (7 days)
p50 587ms·p95 1149msⓘ HTTP response times to Cerebras's status endpoint - measures infrastructure availability, not API inference speed. For TTFT and model-level API status, see the Cerebras API Status section above.
90-day uptime
Incident history
Between 11:00 and 11:15 UTC users experienced partial degradation with Gemma-4-31B. Service has been restored.
GPT-OSS-120B was having partial service disruption between 18:20 UTC to 18:45 UTC. This is resolved now.
Related pages
About Cerebras status
This page tracks the live operational status of Cerebras by Cerebras Systems. We check Cerebras every 5 minutes and record the result. If Cerebras is down or experiencing an outage, it will be reflected here within minutes. Historical uptime data covers the last 90 days.