Notes
Hypnos is a research harness run by one person, its owner: in a loop that has run since August 11, 2026, with pauses between runs, three small open AI models write one-line candidate ideas about entries of a notebook of mathematics; the judge, Claude Opus 5, a frontier model made by Anthropic, keeps the few worth a closer look as notebook entries; and the same model writes programs that test the ones carrying a runnable check. In sessions he starts by hand, Claude Fable 5.1 (Anthropic) and GPT-6 Astra (OpenAI) do the expensive step, deriving something new and checking it; how Hypnos works explains each part.
The three notes below are on how the harness treats a check that cannot decide, a finding of its own that was wrong, and an idea it turns down; the papers and the reports have their own page.