
Sardine AI Incident History
Sardine AI is currently operational with all systems functioning normally.
Incident History
Showing incidents from the last 15 days
Report: "Customer Intelligence tab in US Production Dashboard not showing sessions/transactions"
Last updateThis incident has been resolved.
New sessions are being shown OK, a few older ones might still be missing, they'll be shown shortly.
A fix has been implemented and we are monitoring the results.
The issue has been identified and a fix is being implemented.
We are currently investigating this issue.
Report: "Degraded Performance for EU Dashboard"
Last updateThe issue has been resolved.
We are continuing to monitor for any further issues.
A fix has been implemented and results are being displayed in the Customer Intelligence page. We are monitoring the process until full service is restored.
We are continuing to work on a fix for this issue.
We are continuing to work on a fix for this issue.
The issue has been identified and a fix is being implemented.
We are currently experiencing degraded performance in retrieving the latest customer sessions in the Customer Intelligence page. Root cause has been determined. We are working on resolution.
Report: "Performance degradation on APIs"
Last update**Summary** On 2026-07-15, API requests experienced elevated error rates and increased tail latency \(p99\). Median latency \(p50\) was not affected and there was little effect on 95th percentile latency \(p95\), so most requests performed normally. We began investigating at 14:34 UTC and applied a fix at approximately 19:56 UTC, after which performance returned to normal. **What happened?** Through the affected period, a portion of API requests returned errors, responded with reason code SITO \(Sardine internal timeout\) or timed out rather than completing, with intermittent spikes during which a larger share of requests were affected. **Why did it happen?** A scaling configuration prevented part of our platform from adding capacity as traffic increased. As traffic rose through the day, available capacity was exhausted, causing timeouts and errors for some requests. **What are we doing about this?** We corrected the configuration at **19:56 UTC**, which immediately restored normal performance. We are also making the affected part of the platform more resilient to sudden traffic increases and adding proactive monitoring so we can detect and respond to this class of issue faster in future. We apologize for the disruption.
This incident has been resolved. A Postmortem will be published soon.
We are currently investigating this issue.