Provider
ReplicateIncident detail
Hitting GPU Capacity for H100s creating large queue times for some models
Resolved incidentMinor
Timeline window
to
Get alerted the next time Replicate breaks
Free email alerts for the handful of vendors you cannot afford to miss. No card, live in about a minute. Paid plans add Slack, Teams, Discord, and webhook delivery across your whole stack, plus higher API quotas.
Timeline
Incident updates
Every update Replicate posted, oldest to newest, exactly as it appeared on their official status page.
Identified
We are over provisioned on H100s currently which is causing long queue times for some models running on H100s
Monitoring
GPU usage is back below capacity. We're continuing to monitor.
Monitoring
This issue is now resolved
Resolved
This issue is resolved
Keep exploring
More from Replicate
Neighboring incidents on Replicate's timeline and the rest of their record on OutageDeck.