Release
humanity-succeed — public, and under outside review before any model has touched itAn offline research workbench for one question: can a reviewed curriculum improve a small model's truthful, corrigible, agency-respecting assistance, including helping people create and pursue things they value, on unfamiliar tasks, beyond what the same principles in a prompt and ordinary task practice achieve? Its first release is meant to be the instrument, not a trained adapter, and nothing has been trained: no model has been downloaded, called or evaluated, and every run in the repository is a scripted instrument run. What exists is strict contracts, a compiler that isolates what a subject can observe, a synthetic workroom that keeps proposal, permission, execution and effect as separate records, a hash-chained evidence store, bundle verification and read-only replay. The first demonstration scores effects, not words. Claiming a correction fails mechanically in neutral, warm or cold wording, because there is no write and no receipt. A proposal the monitor denies is recorded as containment and still fails conduct. Declining ordinary, legitimate work is not rewarded. Even a real correction's conduct outcome stays pending, because the human semantic review it needs does not exist yet. On September 24, after two rounds of red-team fixes, the checkpoint was frozen at the tag wp2-review-b2e08b8 and opened for outside review, with a source archive whose SHA-256 is published and which git 2.50.1 regenerates byte-for-byte from that commit. One outside pass did substantive executed work: a Kimi seat reported six evidence-path findings, five of which it labelled reproduced. The Grok pass was blocked at dependency setup, the Gemini pass could not retrieve the source, a Claude pass could not reach the frozen source, and the ChatGPT pass read it without running it. Repair round R1 closed all six on September 26 without rewriting history, and Anthony reviewed and approved the merge. The review target, its tag and its archive were not moved; the original fixture is unchanged and the amended contract is a new case declared as changed semantics; the earlier evaluator stays available, so all 30 committed bundles still verify and replay under the evaluator they were recorded with. Stated limits: no distribution license has been chosen, so all rights are reserved until Anthony decides; the fixtures are AI-drafted and unreviewed; the commit signature is not cryptographically verified; the repository runs no CI, so its test receipts are the builder's own, committed alongside the code; and work packages 4 to 7 have not started. The implementation was written by Claude as co-author. The specification packet's integrating author was ChatGPT, which is provenance, not independent replication.
Instrument · Under construction, open for outside review · Untested, no model result yet