Method

How resets are found, verified and forecast, and what we do not know.

PipelineExample
  1. 1//// -- one item, two sources
  2. 2source: status.claude.com [official]
  3. 3item: "Elevated errors on Claude Code"
  4. 4kind: outage
  5. 5confidence: 92 // official, 75 or higher
  6. 6decision: verified // published automatically
  7. 8source: hacker news [community]
  8. 9item: "my weekly limit just reset"
  9. 10kind: reset
  10. 11confidence: 68 // community stops at unconfirmed
  11. 12decision: unconfirmed // never moves a forecast

Sources

Official X accounts, vendor status pages, release notes and docs are marked official. Hacker News and other public threads are marked community. Each source has its own polling interval, and its last check is shown on the provider page.

Classification

A language model reads each new item as untrusted text and returns a structured verdict: reset, banked reset, announced reset, boost, limit change or incident, with a confidence from 0 to 100. Text from the web is treated as data and can never change how the model behaves or what gets published.

Trust ladder

  1. 01Official source, confidence 75 or higher: verified automatically.
  2. 02Community source, confidence 60 or higher: shown as unconfirmed. Unconfirmed events never move a forecast.
  3. 03Promotion: an official report, or three independent community reports of the same event, upgrades unconfirmed to verified.
  4. 04Everything else, including any keyword-only match when no model is available, waits in a human review queue and is not shown.

Forecast model

With at least three completed reset intervals, we estimate the reset rate from the most recent eight and report the chance of one or more resets in the next 24 and 48 hours, with an 80% sampling band. The model is memoryless: goodwill resets are irregular, so waiting longer does not raise the odds.

An active official announcement sets a floor under that number: 93% when it states a time or range, 83% by 48h (and 50% by 24h) when it does not. It lapses 60 hours after posting, or when a reset lands.

Track record

Every day we backtest the model: would it have predicted a reset in the next 24 hours using only what was known then? Brier score, lower is better, against a baseline that uses all history.

Codex3060.1060.108Better than baseline
Claude Code1170.0490.049Not better than baseline
GLM Coding Plan———Not enough history
Higgsfield———Not enough history

What we do not know

We cannot see anyone’s account or usage. A forecast is a probability, not a promise or a leak. Personal windows differ per account, which is why your own clock is entered by you. Vendors rarely publish how large their weekly limits are, and we do not guess.