How it was worked out
This standard was written by a person and a machine, each correcting the other. Both were wrong more than once, and the record below sets out which was wrong about what. Standards rarely show their working; the discarded versions explain the current one better than the current one can.
“Why does the scale reverse direction?”
The first draft ran L0 to L7. The second extended it to L10. The third replaced the integers with dot notation: 0, then 1.0–1.4 for human-led work, then 2.4–2.0 for machine-led work.
That third version broke the one thing an index cannot break. Inside the 1.x band a higher number meant more machine; inside 2.x a higher number meant more human. The scale reversed halfway along, so nobody could compare 1.4 with 2.1 without a lookup table — and the document contradicted itself in two places without noticing.
The answer was already on a napkin. The originator’s first sketch of the back cover showed a single digit in a box. Every more elaborate scheme was tested against that sketch, and the digit won every time.
“Is the 01 a version, or the index?”
The code was originally a GS1 Digital Link: /01/ followed by the thirteen digits of the ISBN, so that one square on the cover carried both the book’s identity and its record. 01 is the GS1 identifier for a product number, and it would have been 01 on every book ever printed.
The question that settled it was simple: if the person who invented the standard reads 01 as the index, so will everyone. The URL now leads with the number, which is what a reader actually wants.
“The code tells me what a 4 means. It doesn’t need to know which book this is.”
This one reframed the whole system. If the code’s job is to explain the digit rather than look up the work, then it needs no ISBN, no record, and no registration at all.
That is what made adoption free. Any author anywhere prints the mark and the code works, with no account, no database entry and nobody’s permission. Everything that followed — the deleted register, the eleven static pages, the refusal to hold records — follows from that one observation.
“Shouldn’t the index be self-selected?”
The drafting recommendation was a computed one: five questions about conception, structure, drafting, revision and verification, with the number derived from the answers. It removed the boundary arguments no amount of prose can settle, and it was wrong.
A number produced by a form is a number its declarer can disown — the tool said 4 — and a standard resting entirely on honesty cannot afford a gap between the claim and the person making it. The grid was demoted to an optional helper that suggests and never assigns.
It then failed on its own author. Applied to this website the grid returns 7; the honest answer is 6. The mechanism gets the standard’s own site wrong, which settles the argument better than any reasoning could.
“Spellcheckers and grammar checkers do not count.”
Stated flatly, and correct. They have been part of writing since long before generative systems, they correct what the author already wrote, and they put no words in anyone’s mouth. A standard under which every writer alive is a 1 would describe nothing.
Holding that line exposed a defect one level up. Level 1 read “reviewed the finished writing” — which is also what a grammar checker does. Level 1 is now explicitly about substance: facts, continuity, whether the argument holds. Mechanics never reach it.
“If it finds ten timeline errors and I fix them all myself, does that count?”
The hardest question asked of this standard, and the one that produced its sharpest test.
It counts — it is a 1. Nothing the machine produced is in the finished work, but the author did not find those ten errors. Without it the book ships with ten errors in it, and the integrity of the finished work is partly owed to something that was not a person.
Reconciling that with the spellcheck exemption required a second test, which now carries as much weight as the first: did the tool have to understand your work to do its job? A spellchecker would behave identically on a shopping list. Something that catches a contradiction in chapter nine has read your book.
“Does STET suggest a human reviewed it?”
A shortlist of names had reached STET — the proofreader’s mark meaning let it stand. Elegant, publishing-native, proposed with some confidence, and ruled out immediately.
Let it stand is an instruction a human editor gives. At 8 the human only pressed publish; at 9 there is no human at all. The mark would have asserted something the number explicitly denies.
That question became a permanent screening rule: a name must be neutral across all eleven values. It ruled out three candidates at once, and revealed a pattern — the warmest, most bookish names are the least neutral, because their warmth comes from evoking human hands.
Where the machine was too cautious
On trademarks, the drafting advice was gloomy: two research institutes hold the obvious names, the field is crowded, prepare for the answer to be change it.
The reply was one word — Polo. Ralph Lauren, Volkswagen and a mint coexist because trademark classes exist precisely to let unrelated goods share a name. The caution was overdone, and was withdrawn.
Both names did eventually go, but for a different and better reason: a mark containing those letters is weak as well as contested, and describing the thing beats claiming a name for it.
Where the machine got the copy wrong
The mark page read “free, permanent, and nobody needs to know you did.” It was meant to say there is no registration. Read plainly it says something else entirely — that using the mark is something you might want to keep quiet.
On a transparency standard that is the one connotation the page cannot afford. It had spread to three places before it was caught. The rule now is no permission needed, not no witness needed.
The same instinct had produced an earlier fault: routing readers who doubted a declaration straight to that book’s review form. An index that delivers a hostile reader to their target would have become the very thing it exists to prevent.
“The helper is technical. I guessed.”
The five questions were answered against a four-state vocabulary — none, assisted, co-created, generated. Those are the standard’s own analytical terms, and asking a declarer to translate their experience into them was asking them to read the standard first.
If the person who invented the index has to guess, the tool is broken. It was rebuilt as two questions in plain first-person language, the second branching on the first, with every option describing something you either did or did not do.
The lesson was general: a sound analytical model and a good interview are not the same artifact, and one had been shipped as the other.
“Does any of this actually matter?”
Asked directly, halfway through, and it deserved a straight answer rather than encouragement.
The honest one had three parts. The index makes honest declarers visible while dishonest ones stay invisible, so early adopters carry a cost before the network protects anyone. Every previous authenticity panic — photography, synthesisers, sampling, digital images — normalised without ever producing a disclosure standard. And a trade body that calls the practice unethical is not asking how much, so a finer measurement will not move it.
The case for continuing held. The vocabulary works today, with no infrastructure at all, and someone will define this taxonomy whether or not it is someone who actually uses the tools. The conclusion both sides reached: the language is the valuable half, and the printed badge is a much longer bet.
Two names did not survive
The standard was AI² for two drafts. AI2 is a registered trademark of an organisation in the same field.
The obvious fallback was AIAI, matching the original domain. That name is also long established, by a research institute — and a research institute sits in the same trademark class as publishing, which made it the closer collision of the two.
Two names, two prior holders. Not bad luck: every permutation of AI-something is contested, and a mark containing those letters is both harder to clear and weaker once cleared. So the standard stopped claiming a name and started describing itself.
“[AI] {AI} <AI> /AI/”
Four options, offered without comment. Square brackets are the editor’s notation for material that did not come from the author — [sic], [emphasis mine] — which makes them the exact notation for a standard measuring how much of a text did not come from its author.
The follow-up corrected the drafting again. The proposal had been [AI] 4, with the number outside. But the convention encloses the whole interpolation: you write [emphasis mine], never [emphasis] mine. [AI 4] uses the convention properly, and makes the mark one unsplittable token.
It also deleted a section. Earlier drafts needed two notations, a brandmark for display and an ASCII form for metadata. [AI 4] types on any keyboard, so there is one form now — and most use of this standard needs no artwork at all.
“0 is human, 9 is non-human. The intelligence could be anything.”
The last and deepest of them. The standard had been defining its levels in terms of artificial intelligence — a term of art with a shelf life. Nobody says cybernetics now.
The axis is not artificial intelligence. It is human share. Whatever writes prose in twenty years, under whatever name, the question how much of this did a person write? still parses.
That is what makes the permanence promise real rather than aspirational. The level descriptions may be refreshed as the language changes; the human share each number represents may not. A 4 declared in 2026 is still a 4 in 2046, because 4 never meant a technology.
This page has its own history
The first version of it described the changes as though the standard had evolved by itself, crediting nobody. The second overcorrected into a list of the originator’s good questions, which was flattering and equally untrue.
This is the third. It is the only one that mentions the drafting being wrong, which is the part that makes the rest believable.
“Is there ever a case for [AI 5/6]?”
Ruled out, and the reasoning turned out to be less obvious than expected. Seven of the nine boundaries ask whether something happened rather than how much — so a range across them is not a finer measurement but an unanswered question. If artificial intelligence offered judgement on a work, that is a 2, and the 1-ness does not dilute it.
Only two boundaries run by degree: 4/5 and 5/6, both inside the drafting band, which is the one variable on the scale that is genuinely continuous. A work can honestly be sixty per cent rewritten. That is what a rounding rule is for, and every practical index rounds — a shared scale is worth something only while everyone uses the same buckets.
The deciding reason belonged to the reader. A number describes what a reader will encounter, and a reader encounters the whole work. If the first half is a 5 and the second a 6, somebody reading the second half is reading a 6, and a range would conceal precisely what the number exists to disclose.
Who wrote it
The Authorship Index was originated by an author — one person who wrote a book with artificial intelligence and found there was no precise way to say so, only one phrase covering everything from a spellcheck to a machine writing the lot.
Not a company, not a platform, not an institution. That is the whole of it, recorded here and nowhere else: not in the colophon, not on the pages that define the index, and not in the governance commitments. A standard should not depend on who wrote it, and one carrying its author’s name at the foot of every page would be asking to.
The licence binds whoever operates these pages. The text may be copied and mirrored by anyone. If the index outlives the person who started it, it will have worked.
Why 6?
The idea was the author’s. So was every editorial decision — including at least five that reversed the drafting outright: no register, self-selection over computation, the number inside the brackets, human share as the axis, and the whole trademark posture.
The prose is the machine’s. The author did not rewrite it sentence by sentence, which is what a 5 would require. What the author did was conceive the work, direct it, reject a great deal of it, and decide what stood.
That is precisely what 6 describes: a machine wrote it, the author shaped it. Not the most flattering number available, and arrived at by the standard’s own rule — when in doubt, go higher.
A standard whose own site declared a comfortable number would be worth nothing. This is what it actually is.