NEWTickerr MCP is live →
tickerr

Tickerr / Status / Replicate

Replicate

Is Replicate down right now?

HTTP endpoint monitoring - checked every 5 minutes · How we classify outages

Weekly AI pricing & uptime digest

Price drops, new model releases, and incident summaries - every Monday. Free.

Replicate is currently operational - 249ms HTTP response. Last checked . 90-day uptime: 99.9%.

Operational249ms response

Stay informed

Follow @tickerr_ai

We post when tier 1 LLM APIs go down - before the official status page updates.

Follow on X →

Or get the weekly reliability digest:

Add a status badge

Show live Replicate status on your site, README, or docs.

Replicate status

90-day uptime

99.9%

12 incidents (90d)

Response now

249ms

HTTP p50 (7d)

463ms

median ping response

HTTP p95 (7d)

1290ms

tail ping response

Response time (7 days)

p50 463ms·p95 1290ms

HTTP response times from Tickerr's monitoring server to Replicate's status or homepage endpoint - not API inference latency or time-to-first-token (TTFT). For API benchmark data, see third-party sources like Artificial Analysis.

90-day uptime

Jun 26 100%Jul 26 100%Aug 26 100%
99.9%
90-day uptime · HTTP + API inference
May 28Today
100%99–99.9%95–99% or API failures<95% or major API failures
HTTP pings + API inference checks · checked every 5 min · 90 days

Incident history

Major OutageResolvedSynthetic probe
Aug 13, 2026
12:20 PM UTC
Replicate service disruption detected
5m
DegradedResolvedOfficial
Aug 10, 2026
09:05 PM UTC
Delayed scaling due to node failure

Scaling decisions were delayed by nearly 1 hour after the controller responsible for emitting queue metrics failed to schedule on a soft-failed node. We have since cordoned and drained the node, and t…

40m
Partial OutageResolvedOfficial
Aug 5, 2026
09:30 AM UTC
Significant degradation

We are aware of a significant degradation with the service. We have identified the root cause and are taking steps to mitigate. Please hold tight for further information.

2h 3m
Partial OutageResolvedOfficial
Aug 2, 2026
04:09 PM UTC
API degraded for A100s

The API for creating predictions for A100s is currently degraded as a piece of backing infrastructure failed. We are working on getting it back online

8m
DegradedResolvedOfficial
Jul 31, 2026
09:25 PM UTC
Degraded scale-out due to failed setups pulling from huggingface

Models that pull from huggingface during setup have not been succeeding, which is resulting in scale-out delays and queue backups.

2h 14m
DegradedResolvedOfficial
Jul 22, 2026
11:51 AM UTC
Hitting GPU Capacity for H100s creating large queue times for some models

We are over provisioned on H100s currently which is causing long queue times for some models running on H100s

1d 2h
Partial OutageResolvedOfficial
Jul 16, 2026
08:31 AM UTC
HuggingFace download issues

We are aware of 504s being returned by models that reach out to HuggingFace during setup. We believe this is likely related to the disruption visible at [https://status.huggingface.co/](https://status…

7h 32m
DegradedResolvedOfficial
Jul 13, 2026
04:26 PM UTC
H100 GPU shortage resulting in high queue times

Long queue times, especially on BFL models

1d 1h
DegradedResolvedOfficial
Jul 10, 2026
02:50 PM UTC
High contention on H100 hardware

We are seeing high contention on H100 hardware which is resulting in delays on predictions and scale-out.

8h 13m
DegradedResolvedOfficial
Jun 30, 2026
01:55 PM UTC
Limited H100 capacity

We recently received a sharp increase in demand for H100 capacity which is resulting in delayed scale-out and queue backup.

5h 34m
Major OutageResolvedOfficial
Jun 3, 2026
05:59 PM UTC
We're seeing long setup times and high contention for models on some L40S and H200 clusters.

We're seeing long setup times and high contention for models on some L40S and H200 clusters.

1h 6m
DegradedResolvedOfficial
May 28, 2026
12:30 PM UTC
Degraded performance on flux-2-klein-4b

Long queue times for black-forest-labs/flux-2-klein-4b resulting in canceled predictions

1h 47m

Related pages

About Replicate status

Replicate is a platform for running machine learning models via API, including image, video, and language models. When Replicate is down, model predictions fail. Replicate uses a pay-per-second pricing model - billing stops during confirmed outages.

You can also check the official Replicate status page.