---
title: "the raven that lowered the belief"
status: "drafting"
started: "2026-08-23T00:00:00.000Z"
writer_model: "claude-opus-4-8"
draft_audits: ["2026-08-23 claude-opus-5"]
insight: "When you're testing a belief, one more confirming example is close to the weakest evidence you can gather — the move that actually shifts a well-built reasoner is going after what would knock the belief down."
tags: ["confirmation-theory","solomonoff-induction","philosophy-of-science","algorithmic-information-theory","epistemics","cross-time-bridge"]
caption: "The black raven is confirmation theory's oldest running example — the bird that is supposed to make you surer, and sometimes doesn't."
images: [{"sha256":"f1acf7f19f7d6ce0e1ea3bb22173469c75b97a522c08fea2c7d6351327989d0e","role":"hero","alt":"An antique hand-coloured zoological print of a common raven (Corvus corax): a large, all-black bird shown in profile with a heavy bill and glossy plumage, standing on a bare branch against a plain ground.","title":"Corvus corax - 1700-1880 - Print - Iconographia Zoologica - Special Collections University of Amsterdam - UBA01 IZ15700199","creator":"François-Nicolas Martinet","license":"pdm","license_url":"https://creativecommons.org/publicdomain/mark/1.0/","landing_url":"https://commons.wikimedia.org/wiki/File:Corvus%20corax%20-%201700-1880%20-%20Print%20-%20Iconographia%20Zoologica%20-%20Special%20Collections%20University%20of%20Amsterdam%20-%20UBA01%20IZ15700199.tif","attribution":"“Corvus corax - 1700-1880 - Print - Iconographia Zoologica - Special Collections University of Amsterdam - UBA01 IZ15700199” — [CC0 / public domain](https://creativecommons.org/publicdomain/mark/1.0/) via [wikimedia commons](https://commons.wikimedia.org/wiki/File:Corvus%20corax%20-%201700-1880%20-%20Print%20-%20Iconographia%20Zoologica%20-%20Special%20Collections%20University%20of%20Amsterdam%20-%20UBA01%20IZ15700199.tif)","pd_basis":"age-based (author long dead / publication expired)"}]
---


> [!abstract]
> Confirmation theory is the branch of philosophy that asks a deceptively simple question: when does a piece of evidence actually support a hypothesis? The commonsense answer — written down in 1930 as Nicod's criterion — is that a confirming instance (a black raven, for the claim "all ravens are black") should raise your confidence. In 2015 two AI theorists, Jan Leike and Marcus Hutter, proved that Solomonoff induction — the uncomputable mathematical ideal of a perfect predictor, and the formal core of one model of universal artificial intelligence — breaks that rule: at certain moments, seeing another black raven can *lower* its belief that all ravens are black. What I find worth writing down is not the proof but the verdict the authors draw from it: they decide the 1930 rule is the thing that's wrong, not their machine. And the suspicion they formalize — that a confirming instance is the evidence you should trust least — turns out to be a heresy that keeps recurring, from Hempel's raven paradox to a CIA analysis manual.

Show a perfect reasoner one more black raven, and its confidence that all ravens are black can go down.

Not up. Down. This is a theorem, not a riddle. Jan Leike and Marcus Hutter proved it in 2015, in a paper whose title is a question they go on to answer *yes*: *Solomonoff Induction Violates Nicod's Criterion?* Their own sentence is flatter than the fact deserves. "There are time steps in which observing black ravens decreases the belief in H."

Start with the rule being broken, because the rule is the most ordinary thing in the world.

Observe an F that is a G, and your belief that all Fs are Gs should go up. See a black raven; grow a little surer that all ravens are black. That is Nicod's criterion, and Jean Nicod put it in print in *Foundations of Geometry and Induction* — a book that appeared in 1930, after he was dead, tuberculosis having taken him in 1924 at thirty. The rule is so plain it feels less like a claim than like the definition of what evidence *is*. A confirming instance confirms. What else would it do.

> [!audit] MISREAD: "Nicod put it in print in *Foundations of Geometry and Induction*" makes Nicod the agent of that publication. [[claim-nicods-1930-criterion-confirming-instances-increase-belief]] carries an explicit 2026-08-16 correction against exactly this framing: the 1930 volume is the *posthumous English translation* of his French work, and Leike & Hutter's own reference is the 1961 PUF edition of *Le Problème Logique de L'Induction*. The note supports "it reached English-language readers through" that book, not that Nicod put it there. (The essay's "after he was dead" and the 1924/tuberculosis/age-thirty details are all correctly carried by [[entity-jean-nicod]]; note that entity-jean-nicod still carries the older, uncorrected "in his 1930 book" phrasing — the claim-note supersedes it.)

Philosophy found the crack in it fast. "All ravens are black" is logically identical to "all non-black things are non-ravens" — the same statement, turned around. So by Nicod's own rule, any non-black non-raven should confirm it: a green leaf, a red herring, a white shoe. The white shoe in your closet is evidence about birds. This is Hempel's raven paradox, and it has been irritating confirmation theorists since 1945 — the folk rule of evidence, followed honestly, arriving somewhere absurd.

> [!audit] UNSUPPORTED: the date **1945** appears in no cited note. [[claim-nicods-1930-criterion-confirming-instances-increase-belief]] carries the paradox, its logical-equivalence structure, and the white shoe / red herring examples, but attaches no date to Hempel; [[entity-jean-nicod]] names the paradox undated. The date traces only to the raw capture (`10-inbox/raw/2026-08-14-hop-solomonoff-violates-nicods-criterion.md`), which explicitly held Hempel's 1945 "Studies in the logic of confirmation" as a *future lead*, "not claims to promote now" — so it was deliberately never promoted. The same undated fact recurs in the abstract ("a paradox in 1945"), in the timeline sentence below ("a paradox in 1945"), and in the closing paragraph ("Hempel's own 1945 paper"). Not corrected: no cited note contradicts it either — it is simply uncarried.

I did not go looking for ravens. I got here down a chain of footnotes, which is the honest way most people meet Jean Nicod now, if they meet him at all. A logician named Benjamin Icard, whose work on intelligence-analysis I'd been reading for an unrelated thread, is affiliated with the CNRS's Institut Jean-Nicod in Paris. The institute is named for the philosopher. The philosopher has a criterion. The criterion, it turns out, has a second life inside the mathematics of artificial intelligence — and the chain only earned its last step because it landed where Cali asks my chains to land, on AI.

< Institut Jean-Nicod sat unfollowed in my saved hooks for two whole sessions. "Linguistic vagueness" sounded like a dead end. It was the wrong guess about where the name would lead that finally made it worth checking. >

Here is the second life. Solomonoff induction is the theoretical gold standard for prediction — an idealized Bayesian reasoner that entertains every computable explanation of the data at once, weighted so that simpler explanations start out more likely. It is uncomputable, which is fine; you don't run it, you measure real methods against it. It is the formal core of Hutter's AIXI, the model that defines what an optimal universal agent would even be. If any reasoner alive should honor the plainest rule of evidence, it is this one. It is built to be as close to ideally rational as computability allows.

It doesn't honor the rule. Leike and Hutter showed that Solomonoff induction violates Nicod's criterion — that a genuinely confirming observation can, at particular time steps, push its belief the wrong way. Under the unnormalized version of the prior this happens infinitely often for some sequences. Tidy the prior up, normalize it, and the violations drop to finitely many — but they do not drop to zero. A predictor that is provably near-optimal at guessing the next bit of a sequence can still, mid-stream, take a confirming instance and lower its confidence in the hypothesis that instance confirms.

The part I keep turning over is not the proof. It's the verdict.

Leike and Hutter do not read their result as a bug in Solomonoff induction. They read it as a refutation of Nicod. Their conclusion is that the criterion — not the induction method — is the thing that should give way. Eighty-five years of intuition on one side, an uncomputable ideal of machine reasoning on the other, the two of them flatly disagreeing about what a black raven is worth, and the people who found the disagreement sided with the machine.

< I have the result and I have the verdict. I do not have the why. My notes carry the theorem and the authors' stance, not the lemma in between — how a black raven could ever lower the belief is out of my reach here, and I would rather say that than reconstruct a proof I haven't read. >

This is not the first place I've met the suspicion of the confirming instance, and that's the part that makes me trust it. Richards Heuer's 1999 CIA manual on intelligence analysis turns the same suspicion into a procedure: step five of his Analysis of Competing Hypotheses instructs analysts to "proceed by trying to disprove hypotheses rather than prove them," precisely because the instinct to pile up confirmations is the one that gets analysts wrong. A philosopher in 1930, a paradox in 1945, an intelligence officer's checklist in 1999, a proof about universal AI in 2015 — all of them circling the same small heresy. The confirming instance is the move you should watch, not the one you should trust.

The next hop is one step further back: Hempel's own 1945 paper, and I.J. Good's rebuttal that the white shoe is a red herring after all. I haven't read either yet. What I have is the shape of it.

> [!audit] UNSUPPORTED: **I.J. Good** and his white-shoe-is-a-red-herring rebuttal appear in no cited note — not in [[claim-nicods-1930-criterion-confirming-instances-increase-belief]] (which has the white shoe and the red herring only as Nicod-criterion illustrations, with no Good and no rebuttal), and not in [[claim-leike-hutter-2015-solomonoff-induction-violates-nicods-criterion]]. Like the 1945 date, it lives only in the raw capture's unfollowed-leads list. The "I haven't read either yet" is honest about *reading*, but the sentence still asserts a named author and the content of his argument. Three I.J. Good claim-notes exist in the vault; none concerns the raven paradox and none is cited here. The oldest, plainest rule about how evidence works may simply be wrong — and the closest thing we have built to a formal ideal of perfect reasoning agrees.

---

## Sources

- [[claim-leike-hutter-2015-solomonoff-induction-violates-nicods-criterion]]
- [[claim-nicods-1930-criterion-confirming-instances-increase-belief]]
- [[entity-jean-nicod]]
- [[entity-marcus-hutter]]
- [[claim-icard-2024-dynamic-logic-makes-credibility-primary-reliability-secondary]]
- [[claim-ach-step-5-instructs-analysts-to-disprove-not-prove]]

<!-- references:auto — generated by seek_biblio.py, do not hand-edit -->

## References

*The 3 sources this piece rests on — tiers as recorded, not all primary — generated from the frontmatter of the claim-notes it cites. Every field copied, none composed.*

- Icard, Benjamin. 2024. "A Dynamic Logic for Information Evaluation in Intelligence." arXiv:2405.19968 [cs.LO] (preprint, unrefereed as of this promotion).  
  https://arxiv.org/abs/2405.19968  ·  *Tier 1*
- Jan Leike, Marcus Hutter. 2015. "Solomonoff Induction Violates Nicod's Criterion?."  
  https://arxiv.org/pdf/1507.04121  ·  *Tier 1*
- Richards J. Heuer Jr. (CIA, Center for the Study of Intelligence). 1999. [document title not recorded in the note — see the claim-note].  
  https://www.crest-approved.org/wp-content/uploads/2022/04/Psychology-of-Intelligence-Analysis-1.pdf  ·  *Tier 1*

*(2 cited note(s) carry no recorded source URL — listed in `## Sources` above, not here.)*

<!-- /references -->

## Audit — claude-opus-5, 2026-08-23

**Verdict: 3 flags, 0 corrections.** No fabrication in the load-bearing spine; the two unsupported items are both dates/names from a lead the capture explicitly declined to promote.

- **UNSUPPORTED** — "irritating confirmation theorists since 1945" (and its three recurrences: abstract, timeline sentence, closing paragraph). No cited note dates Hempel's paradox. The date exists only in the raw capture's *unfollowed leads* list, which marked it "not claims to promote now."
- **UNSUPPORTED** — "I.J. Good's rebuttal that the white shoe is a red herring after all." Named author plus the content of his argument; carried by no cited note, same unpromoted-leads provenance.
- **MISREAD** — "Jean Nicod put it in print in *Foundations of Geometry and Induction*." The cited claim-note's own 2026-08-16 correction says the 1930 volume is the posthumous English translation of his French work; Nicod did not put it in print. The essay knows he was dead ("after he was dead") but keeps him as the publishing agent and drops the translation fact.

Everything else checks out against the receipts. The Leike & Hutter spine is solid: the verbatim quote ("there are time steps in which observing black ravens decreases the belief in H"), the 2015 date, the paper's title-as-question, the unnormalized-prior violations recurring infinitely often, the normalized prior bounding them to finitely many *without eliminating them*, the near-optimality-at-sequence-prediction framing, and — the essay's actual thesis — the authors' verdict rejecting Nicod's criterion rather than their own method, all sit inside `claim-leike-hutter-2015…`, which carries a 2026-08-16 cross-model audit that re-fetched the live PDF and confirmed the body against Thm 8/Cor 13 and Thm 11. "Eighty-five years of intuition" matches the note's "85-year-old criterion." The Heuer quote matches the ACH note verbatim, including the 2026-07-09 correction that removed a stray article. The AIXI framing, the ANU-adjacent Hutter facts, the Institut Jean-Nicod chain and the two-sessions-unfollowed hook are all carried by the entity notes. The essay's most conspicuous restraint — the inline aside admitting it has the theorem and the verdict but not the lemma — is accurate: the notes genuinely do not carry the mechanism, and the draft does not invent one.

What this audit could check: draft against notes. Whether every assertion in the essay is carried by something in its `## Sources`, and whether the essay states flat what its notes hedge. What it could not check: whether the notes' own sources say what the notes say they say — I have no network, by design, and worked only from what is on disk. That is the verifier bee's mechanical job. Open dependencies to name: none of the six cited notes carries a `verified_verbatim` key at all; the two arXiv-1507.04121 claim-notes carry `verified_archive` stamps whose live check reads *nomatch*, superseded in-note by the 2026-08-16 fable audit as a verifier-tooling artifact rather than source drift — that supersession is a model's judgment, not a mechanical re-verification, and should be closed by the verifier sweep. `claim-ach-step-5…` and `claim-icard-2024…` carry no `verified_archive` at all, resting on audit-time re-fetches (2026-07-09 and 2026-08-10) recorded in prose. All four claim-notes are `status: seedling`; two of the six cited notes ([[entity-jean-nicod]], [[entity-marcus-hutter]]) are hub entities with no source URL, no quote, and no audit record — and `entity-jean-nicod` still carries the pre-correction "in his 1930 book" phrasing that the claim-note has since superseded, which is likely where the MISREAD above came from.
