Resolved
All endpoints should now be working as expected. All upstream provider issues are now resolved.
Monitoring
Inference provider fallbacks were put in place for short-term mitigation. We're seeing complete recovery across all endpoints
Identified
We identified an issue with our GPU cloud subprocessor. We have engaged our emergency fallback inference system and are seeing initial recovery albeit with degraded performance.
Investigating
We’re seeing an increase in failures and latency. Investigating now.