{
  "meta": {
    "version": "v1",
    "pricing": {
      "public": {
        "label": "Public",
        "description": "Read-only API access for lightweight status checks and public integrations."
      },
      "premium": {
        "label": "Premium",
        "description": "API keys with higher hourly quotas, plus Slack, Discord, webhook, and email outage alerts across your vendor stack."
      }
    },
    "generatedAt": "2026-07-24T10:17:20.208Z"
  },
  "data": {
    "id": "incident_statuspage_buildkite_cwsbddzngdlq",
    "slug": "buildkite-increased-latency-and-error-rates-2026-05-26",
    "title": "Increased latency and error rates",
    "summary": "Increased latency and error rates",
    "status": "resolved",
    "severity": "minor",
    "startedAt": "2026-05-26T09:56:49.282+00:00",
    "updatedAt": "2026-05-28T04:21:20.334+00:00",
    "resolvedAt": "2026-05-26T10:38:18.23+00:00",
    "provider": {
      "slug": "buildkite",
      "name": "Buildkite"
    },
    "affectedServices": [],
    "links": {
      "html": "/incidents/buildkite-increased-latency-and-error-rates-2026-05-26",
      "api": "/api/v1/incidents/buildkite-increased-latency-and-error-rates-2026-05-26",
      "providerHtml": "/providers/buildkite"
    },
    "impactSummary": "Buildkite reported a none event for the affected tracked services.",
    "source": {
      "id": "source_buildkite_status",
      "kind": "official_status_page",
      "name": "Buildkite Status",
      "checkedAt": "2026-07-23T12:00:00Z",
      "officialUrl": "https://www.buildkitestatus.com",
      "statusPageUrl": "https://www.buildkitestatus.com"
    },
    "updates": [
      {
        "id": "update_statuspage_buildkite_cwsbddzngdlq_1k954h02p2dc",
        "status": "identified",
        "body": "We're observing increased latency and error rates for a subset of our customers. We're currently remediating and will provide status updates as they become available.",
        "createdAt": "2026-05-26T09:56:49.397+00:00"
      },
      {
        "id": "update_statuspage_buildkite_cwsbddzngdlq_j7p8tjwbfysf",
        "status": "monitoring",
        "body": "We've identified the problem and have completed the remediation steps, we are now monitoring as service resumes.",
        "createdAt": "2026-05-26T10:08:16.507+00:00"
      },
      {
        "id": "update_statuspage_buildkite_cwsbddzngdlq_c1d1q9gs0822",
        "status": "monitoring",
        "body": "We see processing time for all affected services has returned to normal as of 20 minutes ago.",
        "createdAt": "2026-05-26T10:37:43.407+00:00"
      },
      {
        "id": "update_statuspage_buildkite_cwsbddzngdlq_zpxt3cm33mt7",
        "status": "resolved",
        "body": "We think the impact from the issue is over.",
        "createdAt": "2026-05-26T10:38:18.23+00:00"
      },
      {
        "id": "update_statuspage_buildkite_cwsbddzngdlq_sgkwbxjhk7kt",
        "status": "resolved",
        "body": "## Service Impact\n\nA subset of customers experienced elevated latency in notification delivery.\n\n## Incident Summary\n\nWhile migrating a subset of our background processing services to Amazon EKS, we encountered an issue with delivery of internal metrics. The discovered issue did not impact performance or availability, but would have impaired our ability to detect such problems if they occurred.\n\nOut of an abundance of caution we decided to revert the migration, and moved those services back to the original infrastructure on AWS Fargate.\n\nWhen migrating to EKS, we scale down and disable automatic scaling on Fargate. This allows us to quickly migrate back by scaling up Fargate. When we moved the workloads back to Fargate to restore internal metrics, we missed the step to re-enable autoscaling. As a result, the affected services did not have sufficient capacity and could not keep up with incoming work.\n\nWe re-enabled autoscaling promptly once the problem was discovered, and provisioned extra capacity for customers where a backlog of work had accumulated.\n\nBetween 09:17 and 10:17 UTC, a small subset of our customers were impacted. Individual customers experienced a limited outage of notification services, which lasted between 35 and 58 minutes within this window, if there was any impact at all. The migration is performed in small batches, so not all customers experienced this incident.\n\n## Changes we're making\n\n* We are simplifying the runbook used to rollback migrations in the event of incidents.\n* We are adding more verification steps to the migration process.",
        "createdAt": "2026-05-28T04:20:36.259+00:00"
      }
    ],
    "access": {
      "plan": "public",
      "keyed": false
    }
  }
}