The Ostracon · the working estate⌂ the tomb
A worked example, offered for inspection · prepared for readers of the evidentiary question
the full record: one level up ↑
Mind the Gap · the making record · 10 May – 9 June 2026

Contemporaneous documentation of human creative control in an AI-assisted work: a complete instance

The Office’s Part 2 Report (Jan. 2025) settled that AI-assisted works are registrable where a human exercises meaningful creative control over the expressive elements; Thaler v Perlmutter (No. 25-449, cert. denied Mar. 2, 2026) fixed the human-authorship floor while leaving open how much human involvement suffices. The live question — posed directly in Allen v Perlmutter (D. Colo., pending) — is evidentiary: what does demonstrating control look like in practice? As of July 2026, cross-motions for summary judgment in Allen are fully briefed and await decision.
This record is one worked answer, preserved rather than argued: the complete working corpus of a 59,865-word trade book produced with AI assistance over thirty days — every correction logged, every voice-preservation decision visible, every version difference tracked, captured at the moment of creation and not reconstructed for any proceeding. Where A Single Piece of American Cheese registered on 35 documented rounds of iterative modification, and the Zarya decision preserved authorship at the selection-coordination-arrangement layer, this record documents that layer continuously, at corpus scale, with the human and machine contributions separable throughout.
And the posture here is not Allen’s. There, the expressive output is wholly machine-generated and the human case rests on the prompts; the Office’s briefed reply is that prompts are instructions, that the system determines the final expression, and that effort and iteration do not substitute for authorship. Here the contributions are separable at every step: 55,693 words typed by the author, 1,621 directive turns steering selection, coordination and arrangement across 86 versions, a human reviewer’s 37 page-keyed items carried in as tracked changes — every acceptance the author’s, in Word, by hand. The claim is never the volume of the iteration. The claim is a documented human hand at the layer Zarya preserved, held continuously, at corpus scale.
The author typed 1.5% of the corpus — and made 100% of the decisions.
Of the 1,621 directive turns, 346 were ratifications and 177 were corrections or outright refusals — roughly one human turn in three was judgement passed on machine output rather than new instruction. Every turn was classified by hand after an automated pass was tested against a blind sample and rejected as inadequate; the labels validated 100/100 against two held-out samples, and the count reconciles to the register exactly.
The record holds 52 documented rejection events — a floor twice over: the count covers the chat channel only, hand-validated from every candidate, and the author’s accept-or-reject rulings on tracked changes in Word are a second channel the transcripts cannot see. Among them: the machine’s proposed ending, killed for the author’s own typed closing lines — the book’s last words are his.
And the typed words that survive are the load-bearing ones. Of the author’s 55,693 typed words, 2,155 — 3.9% — survive verbatim or near-verbatim in the print master, and the survivors are the title, the dedication, the epigraph, the disclosure, the aphorism the book is built on — each surviving passage mapped to its printed page. Small by volume; the spine of the book by weight. Survival is not proof of origination direction — dictation and quoting-back-while-editing both count, and both are authorial acts.
The silent channel is measured too. 120 of the 257 retained manuscript files still carry the machine’s tracked changes frozen in situ — the proposals preserved beside the author’s decided resaves, often minutes apart, the channel unbroken from the first version to the print files. Of fifty measured edit rounds, thirty-four were accepted whole, fourteen mixed by hand, and two rejected entire — in Word, leaving no chat trace. At one of those, the author then typed the back-cover page himself: thirty-three tracked insertions under his own name.
The record, by evidentiary function
The author’s directions, reversals and refusals, day-dated — 1,621 directive turns across thirty days: the selection, coordination and arrangement record, kept as it occurred.
The division of labour, seat by seat: eighteen named production seats — the book’s printed fourteen, reconciled on the ledger — and one human reviewer. Which contribution originated where, in which versions, under whose direction. Human and machine work separable, not blended.
Aggregates hide the machine. Pick one sentence, one chapter, one day — and walk it through every hand it passed through: a reviewer’s question, a machine audit, the author’s own forty-five words, a tracked change accepted by hand, a printed anchor. Twelve chains, each ending at the printed page or a verified deliberate absence.
86 manuscript versions; a 3,792,153-word working corpus, disk-measured; each published figure stating the instrument it was measured with. The arithmetic is checkable and intended to be checked.
The declarations against interest: the record was assembled by the apparatus it describes, its workers graded themselves, and a single author ruled on his own evidence. Stated in full, unsoftened, before anything else is weighed.
How the record was kept: the union method for worker histories, the enumerated/sampled distinction on every count, and the three-grade confidence marking (verified / inferred / assumed) applied throughout.
The wider landscape · the live cases this record sits among
Two questions are being argued at once. This record speaks to the first — whether AI-assisted output is authored; the second, the input question, is the adjacent terrain the same readers cross. All public record, cited for orientation, not as claim.
Output & authorship — the question this record answers
Thaler v Perlmutter
No. 25-449 · SCOTUS cert. denied 2 Mar 2026
Fixes the human-authorship floor, while leaving open how much human involvement suffices — the gap this record is offered into.
Allen v Perlmutter
D. Colo. · cross-motions for summary judgment fully briefed · decision pending
Asks the live question directly: can iterative, creative prompting (some 600+ prompts) amount to human authorship? The closest active case to the question this record answers — and the record’s posture is deliberately not Allen’s: the contributions here are separable throughout, the typed words counted, the acceptances made by hand. Closer to the layer Zarya preserved than to the prompts Allen defends.
A Single Piece of American Cheese
USCO registration · Jan 2025
Registered on 35 documented rounds of iterative modification — authorship in the selection, coordination and arrangement, not in the prompts. The nearest registered precedent.
Zarya of the Dawn
USCO decision · Feb 2023
Text and the selection-coordination-arrangement of the images registered; the AI-generated images themselves not. Establishes the curatorial-authorship layer this record documents at scale.
Burrow-Giles Lithographic Co. v Sarony
111 U.S. 53 · 1884
The old foundation, cited in every modern decision: authorship requires a human originator who fixes an intellectual conception in tangible form. The doctrine is old; the question is new.
Training data & input — the adjacent landscape
Bartz v Anthropic
N.D. Cal. · 2025
Training on lawfully acquired books held “spectacularly transformative”; retaining pirated copies, not fair use. A record settlement followed.
Kadrey v Meta
N.D. Cal. · 2025
Summary judgment for Meta on a different fair-use analysis — likely circuit-split material in time.
New York Times v OpenAI / Microsoft
S.D.N.Y. · filed 2023
Focus on regurgitation — whether models reproduce training text in their output. Consolidated, pending.
Thomson Reuters v Ross Intelligence
training on legal-research data
Held NOT fair use — the counterpoint that the courts are not uniform on the input question.
The doctrinal frame is the US Copyright Office’s AI reports: Part 1 (Jul 2024, digital replicas), Part 2 (Jan 2025, copyrightability of AI output — the source of the meaningful-creative-control test), and Part 3 (May 2025, training data).
The full record, reordered for counsel → — the same evidence, arranged for the questions your field is arguing.
What is not claimed. No legal theory is advanced here, and no conclusion about registrability is asserted. This is a documented instance of the practice the doctrine is attempting to describe. Its weight, if any, is for the reader to determine.
Notes for the careful reader. The book’s printed pages put the worker count at fourteen; the post-publication census resolved eighteen production seats, and the correction is dated on the ledger — the record repairs itself in public rather than quietly. Likewise the versions: the printed figure of record is 86, V0 to V85; on disk the integer chain closes at V84 and hands to the production sub-line, and two version labels — V43 and V85 — carried working meaning without ever surviving as files. The reconciliation is on the ledger. Every book figure on this estate is counted on one frozen file, the Ed1.05 print master, named in Governance; the author’s working copy drifts, the record’s counts do not. And the record does not timestamp itself alone: Nielsen allocated the book’s ISBNs at 08:06 UTC on 11 May 2026 — the morning after genesis — the production ladder is stamped by the publishing platform to the minute, and the Internet Archive holds this site’s first snapshot. The independent anchors are inventoried in the record.
Further materials, on request. A methodology assessment and a case-for / case-against charter exist at reviewing depth; the underlying evidence subset is available by consent. Correspondence: hello@paulroebuck.co.uk.
Paul Roebuck · Mind the Gap (2026) · SHaDS™, SHaDSy™, Additional Intelligence™ are claimed marks.
the record is self-contained · no first-party analytics or behavioural tracking on this page; hosting logs are operational only
SHaDS™ · SHaDSy™ · Additional Intelligence™ · Paul Roebuck IP, 2026.
the Ostracon team  p·18/55