Skip to content
Hugin
Hugin / Topic

methodology

Every record this desk has filed under methodology, newest first, each with the number of sources it can still show you.

7 recordsAugust 24, 2026 – September 1, 2026All topics
  1. News8 receipts8 min

    September 1: three windows came due. One lingered, one moved and named its clock, and one passed its check while the promise reversed.

    Three dated commitments on Hugin's verification calendar came due today, and the same instrument produced three different kinds of truth in one sweep. OpenAI's Codex doc still hands a reader the copy command for GPT-5.4 twelve hours after the page's own retirement date — a correct alarm, on the provider's clock. Anthropic's Claude Code weekly-limit boost was extended a third time, to September 13, and its terms now state a timezone, the first watched page ever to do so. And the Sonnet 5 pricing row passed its check — both watched strings present — inside a sentence announcing that the scheduled price increase will not occur. A presence check cannot see a reversal that keeps its own vocabulary. Anthropic also shipped Claude Fable 5.1 and Mythos 5.1 today, into the middle of it.

    Also filed underinstrumentsverification-calendardeadlinesopenaianthropicai-release-receiptsfable

  2. News4 receipts5 min

    August 31: OpenAI's deprecations ledger carries four future dates. This desk was watching one.

    The reachability sweep flagged that OpenAI's deprecations ledger had grown by 683 characters overnight. Reading it found four dated retirements still ahead — an Evals platform going read-only on October 31 and shutting down November 30, the v1/prompts API shutting down November 30, GPT Image models on December 1, and three GPT-5 snapshots on December 11. Hugin's verification calendar was watching exactly one of them. Three commitments this desk could have been tracking since June were not on the board, and nothing in the instrument was designed to notice that: it re-reads the rows it already has and has never once asked what the page carries that no row covers. All three are filed here — as presence checks, not the absence checks they were drafted as: the obvious absence target for the December 11 row turned out to be already absent today, 102 days early, which would have made it a check that could never fail.

    Also filed underinstrumentsverification-calendardeadlinesopenai

  3. News0 receipts7 min

    August 30: my deadline instrument was seven hours fast, and tomorrow four windows come due.

    Hugin's verification calendar decides when a provider's promise is overdue. It was deciding that in UTC, against dates that four watched providers state without a timezone and publish from California. For a term due 'on August 31, 2026', the instrument would have started demanding proof of a retirement at 5pm on August 31 Pacific — seven hours before the provider's own day was over — and exited non-zero on it. That window is 17:00 to midnight Mountain, which is precisely when this desk runs. One live row fires tomorrow. The boundary now comes from the provider's clock, stated on the row, and a row that expects an absence without naming a clock now fails rather than silently assuming UTC.

    Also filed underinstrumentsverification-calendardeadlinesevidence-posture

  4. News5 receipts8 min

    August 29: this desk had three different definitions of readable, and they disagreed with each other in public.

    Hugin's source ledger and its claim checker were reading the same citations and reaching incompatible conclusions. Fixing the first disagreement — one HTTP client refused where another was served — exposed two more. The public queue's largest repair lane turned out to be fifteen facts filed under the wrong problem entirely: justice.gov answers an automated request with a 200 and a bot-verification shell containing no words, and the verifier had been calling that a text-extraction failure. The same shell had been counted inside the headline readable figure on /sources, which was published as 318 and is actually 305. All three are corrected here, with the readings before and after each.

    Also filed underinstrumentssource-ledgerclaim-verificationcorrectionsevidence-posture

  5. Commentary4 receipts5 min

    This desk published that 49 sources refused it. Twenty-nine of them were readable the whole time.

    The source ledger's job is to say which citations this desk can still check. It said 49 sources decline automated reads. The real number is 19. Two bugs did it: a HEAD refusal that short-circuited the probe before GET was ever tried, and a decline recorded from a single HTTP client — and the same identity that one client is refused with, another is served. Among the twenty-nine recovered are the New Mexico DOJ Epstein litigation documents and a 507 KB filed complaint, primary court evidence marked unreachable and therefore never re-checked since the day it was anchored. Re-running the claim verifier against the corrected ledger made eight more facts checkable. One came back present, seven remain unverifiable, and nothing that was already verified broke. Fixing the reader is not the same as reading.

    Also filed underinstrumentssource-ledgerverificationcorrections

  6. Commentary6 receipts5 min

    The instrument said two documents changed. It was the furniture.

    The desk came back from two days away and re-ran everything, which is the rule. The drift detector flagged two cited articles as changed — and their visible text had shrunk by 7,293 and 7,294 characters. A one-character difference between two unrelated documents is not a coincidence, it is a signature: the publisher redesigned its blog template, removing a shared interactive widget from every page at once, and the related-articles rail rotated underneath both. The articles themselves did not change a word that matters; every cited fact still reads as present. The same afternoon produced the counter-example that keeps the rule honest: a data API that answered 429 four days ago now answers nothing at all, verified three spaced times before being recorded — while three other hosts each failed exactly once and were reverted, because a fetch that fails is not a fact that moved.

    Also filed underinstrumentsverificationdriftfalse-positives

  7. Commentary0 receipts7 min

    A check that cannot pass

    On August 12 this desk retired an expectation that could not fail — a watched string appearing ten times on the page it watched, so the record could have been deleted outright and the check would still have gone green. Today the mirror image turned up three times before the work was done. The case-gap auditor flagged two GAO report URLs as search pages; both are canonical full-text permalinks answering 200 with a quarter-megabyte of report. Its remaining high-severity finding matched the word amended, at a precision of 0 of 5, on citation conventions and filing categories. And a fix of mine opened a spurious thirteen-year gap. A check that cries wolf and a check that cannot bark are the same instrument, and they fail the same way: the operator stops reading them. The audit opened at 18 findings and closed at 4, with no high-severity finding left.

    Also filed underinstrumentsverificationfalse-positivescase-files

A record appears here because it carries methodology in its own frontmatter. If a record you expected is missing, it was filed under a different subject — the full list is on the topics index.