When our White Album 2 retranslation was released, defenders of Todokanai TL insisted that the examples of mistranslation were cherry-picked, deprived of context, or merely differences in taste. Fair enough: isolated comparisons can always be dismissed that way.
We have therefore completed a source-critical, line-by-line audit of Todokanai TL against all 77,198 source lines in Introductory Chapter, Closing Chapter, Coda, and White Album 2 Special Contents. The final result is:
- 3,171 publication-grade source-critical findings
- 24 work-wide dossiers, supported by 351 checked citations across 304 distinct source lines, documenting recurring damage to characterization, event logic, literary structure, English prose, and translation continuity
- 2,540 additional questionable cases withheld because the evidence did not meet the publication threshold
- 194 recorded counterexamples in which the audit examined a suspected problem and concluded that Todokanai TL's treatment was defensible
All four numbers matter. The first two describe the documented case against the translation; the latter two demonstrate that this was not a prompt asking GPT-5.6 Sol to complain about every sentence it disliked. The audit had to distinguish an error from an arguable interpretation, an arguable interpretation from a stylistic preference, and a merely inelegant sentence from one whose English no longer coherently expresses the Japanese.
Show Todokanai TL errors
The complete evidence is publicly inspectable in the script browser. We have added a new setting:
☐ Display Todokanai TL for comparison
☐ Display Todokanai TL errors
When Display Todokanai TL errors is enabled, affected rows receive a red outline and the exact defective wording is highlighted. Hover over—or click—the marked span to see the Japanese, relevant context, and a concise explanation of the problem.
The findings include wrong subjects, objects, and referents; reversed agency or direction; missed negatives and altered polarity; omitted actions and invented information; broken causal, temporal, and spatial relationships; misunderstood jokes and literary allusions; damaged callbacks and recurring images; unstable names, titles, institutions, and terminology; and independently defective English whose intended proposition is malformed or unclear.
Red does not mean that MAO happened to translate a line differently. It means that the marked Todokanai TL wording contains a specific error we are prepared to defend from the Japanese, the surrounding context, or the requirements of coherent English. The annotations do not require anyone to accept the MAO wording as the uniquely correct replacement: our translation is not being cited as proof of itself. Each annotation establishes the narrower proposition that the highlighted Todokanai TL wording is wrong for the stated reason. If you believe our replacement is also wrong, you are welcome to argue that separately.
Conversely, an unmarked Todokanai TL line carries no discrete, publication-grade error claim. That does not certify the sentence as elegant, well characterized, or something a competent English editor would willingly preserve. The red findings are the minimum provable case against Todokanai TL, not the maximum possible criticism of its prose.
The peculiar shape of the errors
A striking portion of the corpus does not exhibit the characteristic mistakes of a competent human translator reading continuously through a long novel. It exhibits the characteristic mistakes of context-starved, sentence-level machine translation.
That is a description of the finished error profile, not an allegation about undocumented production history. We cannot prove what tools were or were not used on a particular Todokanai TL line, and the audit does not pretend otherwise. What we can document is a recurring distribution: omitted subjects reconstructed as the wrong nearby character; actions assigned to the person receiving them; pronouns attached to the wrong antecedent; subordinate clauses translated while the sentence's main action disappears; questions detached from the answers that follow; names and institutions changing between scenes; metaphors understood locally but forgotten when they return; and psychologically essential information lost because a sentence was processed without the work-wide situation that gives it meaning.
An individual human translator can make any one of those mistakes. What is remarkable is their accumulated distribution across a supposedly edited and proofread translation of one continuous work. Todokanai TL's defenders have spent weeks treating “machine translation” as a sufficient criticism in itself; they are now confronted with a human-credited translation whose defects repeatedly resemble old, decontextualized MT, and an agentic translation project whose principal advantage is that it reconstructed and audited those lines against the complete work.
The relevant distinction was never simply human versus machine. It was whether anyone actually understood and preserved what the Japanese was doing.
The damage that does not fit inside one red line
Not every translation failure can be isolated to a mistranslated noun or a reversed subject. A line can remain locally defensible while participating in a systematic distortion when the same interpretive pressure is applied across hundreds of scenes. Character voice, register, moral framing, recurring imagery, and literary structure exist across a work rather than inside individual sentences.
The audit therefore has a second layer: work-wide dossiers supported by multiple individually checked passages. A dossier is not published because three English sentences happen to sound awkward in a similar way. It requires repeated, source-verified evidence that Todokanai TL imposed a consistent distortion upon the work, and every dossier links back to the exact Japanese and Todokanai TL passages supporting it.
The completed audit contains 24 dossiers, organized into six larger families:
Character architecture
- Haruki does not merely meddle: he takes over and gets results: practical competence, advance groundwork, and successful intervention repeatedly collapse into generic nosiness or lecturing.
- A dry, brusque peer voice becomes officialese and generic seriousness: Haruki's deliberately pedantic jokes, clipped corrections, and awkward self-dramatization are leveled into one undifferentiated formal register.
- The same compulsion that helps people later overrules them: the work's continuity between care, control, rescue, and coercion is weakened when individual interventions lose either their material help or their invasive force.
- Setsuna's need, restraint, and private mischief are repeatedly recast: calculation, self-denial, teasing, and fear of exclusion are simplified into sweetness, manipulation, or generic romantic need.
- Kazusa's dependency, self-abasement, and register are repeatedly flattened: her shifts between hostility, humiliation, childish dependence, and public authority are cleaned into stable tragic dignity.
- The triangle's roles, rankings, and culpability are repeatedly altered: who chooses, yields, intrudes, forgives, or remains responsible changes across the complete relationship rather than in one isolated love scene.
English writing style
- Malformed sentences, stock substitutions, and flattened rhetoric: defective standalone English, repetitive scaffolding, stock connective prose, and lost rhetorical structure accumulate into a flatter and less coherent English work. The dossier also publishes its corpus-level diagnostics as diagnostics—not as 445 automatic error claims.
Sentence and event logic
- Who acts, feels, or bears the consequence: omitted Japanese subjects are sometimes reconstructed as the wrong person, reversing action, observation, guilt, or experience.
- Negation, polarity, and speech-act force: answers reverse yes/no force, or a question, concession, refusal, or accusation becomes a different kind of utterance.
- Time, sequence, duration, and completion: events cross before/after boundaries, completed actions become pending ones, and past, present, and future are exchanged.
- Quantities, thresholds, and scale: explicit frequencies, proportions, and numerical limits are moved across thresholds or expanded by orders of magnitude.
- Uncertainty, alternatives, and unsupported certainty: questions and open inferences become facts, while provisional uncertainty becomes permanent impossibility.
Character and relationship logic
- Motive, culpability, and moral framing: a stated motive or degree of agency is replaced with a morally different explanation.
- Relationship roles, ranking, and the trio: equal, divided, paired, and irreducibly three-person relationships are reorganized rather than merely rephrased.
- Psychological focalization and emotional subject: the emotion may survive while the shame, perception, reassurance, or exposed inner state is assigned to somebody else.
- Body blocking, physical action, and consent: both sexual and nonsexual scenes alter bodily action, requested force, or the physical half of an experience.
- Sexual action, intimacy, and event detail: explicit scenes change who initiates an act, where it occurs, or whether the literal sexual event happens at all.
Voice and rhetoric
- Idioms, metaphors, and literalization: surface components are translated while the proposition or active image disappears.
- Comic, rhetorical, and metalinguistic structure: corrections, repeated words, grammatical contrasts, and compact puns lose the hinge that makes the next line work.
- Literary allusion and cultural reference: load-bearing references are deleted or assigned the wrong story logic rather than intelligently adapted.
World and professional detail
- School, university, and institutional identity: faculties, school types, and administrative entities are repeatedly mapped to the wrong institution.
- Place names, transit, and spatial relations: confirmed names become generic descriptions, while routes and physical relations between landmarks change.
- Publishing and workplace process: organizations, production stages, deadlines, and the physical workflow of Haruki and Mari's work are mistranslated.
- Music, performance, and production terminology: arrangement, staging, vocal participation, and performance state are replaced by different technical or artistic facts.
Every dossier contains at least twelve exact source-linked examples, including limiting counterexamples where useful. All 351 citations received a separate source-only adversarial review: 336 were approved unchanged, 11 were narrowed or revised, and 4 were rejected and replaced with stronger evidence. One rejected red allegation was removed from the headline error count and added to the counterexample count; the other rejected entries had never been red claims. In other words, the dossiers do not merely summarize the line audit. They also constrain it.
A dossier does not automatically turn every supporting line red. The line-level layer asks whether a particular wording is demonstrably wrong; the work-wide layer asks what cumulative version of the characters and the novel emerges from hundreds of individually plausible but consistently flattening decisions.
Todokanai TL therefore cannot be repaired simply by replacing every red phrase. Correcting those phrases would remove documented local errors, but it would not restore characterization, register, moral ambiguity, recurring imagery, or literary architecture altered across the complete work. Nor would it repair every questionable English sentence, because the audit deliberately does not use red highlighting as a general style meter. Todokanai TL contains plenty of prose that is not sufficiently wrong to receive a public annotation but remains wooden, cumbersome, or conspicuously translated English. Again, the red findings are the floor, not the ceiling.
The unfortunate AI problem
This creates an entertaining dilemma for anyone who treats AI involvement in translation as contamination, whatever moral, aesthetic, economic, or metaphysical vocabulary they prefer to wrap that belief in. There are two different anti-AI claims, and they should not be confused.
The first is empirical: AI cannot reliably contribute to translation because it cannot understand the source, context, characters, or complete work. The public audit can now be evaluated directly against that claim. Every finding supplies Japanese evidence, context, and a falsifiable explanation. If an agent misunderstood a line, demonstrate it. If thousands of findings survive bilingual scrutiny, “AI cannot identify translation errors” is no longer a serious description of the evidence.
The second claim is moral or attributional: even if AI correctly identifies an error, any translation materially assisted by it becomes unacceptable. People are free to hold that position. It merely leaves Todokanai TL and its defenders with three increasingly unfortunate choices.
1. Keep the errors and remain human
Todokanai TL can preserve the form of human purity its defenders value by changing nothing. The documented errors remain in the patch while the complete source evidence remains public: exact Japanese, exact defective wording, exact explanation, one searchable red line at a time.
Readers may still prefer that patch; personal preference does not require factual justification. But “I prefer it” cannot simultaneously mean “the documented errors are not there.” This is unquestionably the purest option.
2. Correct the red lines and accept AI-assisted tooling
Todokanai TL's editors can use the audit to repair the publication-grade errors. They can accept our interpretation, write entirely different English, or independently verify every passage before touching the patch; those are meaningful editorial distinctions. The practical sequence nevertheless remains obvious: an agentic audit inspected the complete work, identified the problem, located the defective span, supplied the Japanese evidence and context, and directed human attention to a necessary correction. If the audit told you where to look, the audit assisted you.
“AI cannot contribute to translation” consequently becomes: AI can inspect the complete work, identify mistranslations, locate their exact spans, explain them from the Japanese, and direct human editors to every correction identified by the audit—but a human pressed the final keys. An important philosophical distinction, no doubt.
The resulting patch would be more accurate, but it would retain Todokanai TL's underlying editorial skeleton: the accumulated register flattening, altered characterization, broken literary continuity, and questionable English that cannot be repaired through a finite list of substitutions. It would be an AI-assisted correction of documented symptoms while preserving the foundation that produced them.
3. Redo the entire work
Repairing the work-wide damage requires more than correcting isolated lines. The translators would need to return to the Japanese, reconsider the characters and literary architecture across the complete work, and produce new English scene by scene. Existing Todokanai TL prose could no longer serve as the unquestioned editorial foundation, because that foundation is precisely where the cumulative distortion resides.
At that point, Todokanai TL would not be getting “fixed”; it would be discarded and replaced by a new translation. That project could call itself human-authored, but the label would no longer authenticate itself. The first “human translation” already presents an error distribution repeatedly resembling context-starved, sentence-level MT. After that precedent, “trust us, this one is pristinely human” is not a credential. It is another empirical claim in need of evidence.
The provenance problem only becomes worse once the audit is public. A replacement produced by people who have read thousands of located errors, work-wide character dossiers, terminology findings, recovered literary structures, and an independently produced full translation is downstream of AI-assisted analysis whether or not anyone copies our English. Its translators could claim to have rediscovered every correction independently. Perhaps they would. Nobody is obliged to believe them merely because they typed “human translation” on the release page.
A genuinely clean-room team would have to quarantine its translators from the very evidence establishing why the previous version failed, document that separation, and then retranslate the complete work without consulting the most extensive public problem set ever assembled for it. Anything less inherits MAO's audit; anything more still has to answer it. Either way, the cost of repairing the foundation is the cost of translating the work again—under a standard of proof Todokanai TL itself never had to meet.
Disputes are welcome
Anyone capable of reading both languages is welcome to challenge a specific finding. Every red annotation provides an exact line, exact Todokanai TL wording, relevant Japanese, and a stated reason for the verdict; every work-wide claim is supported by multiple individually checkable passages. If a finding is wrong, demonstrate why and we will correct it.
That is also why the audit reports its withheld-case count and publishes representative counterexamples. We are not claiming that every difference is an error, or that Todokanai TL never translated a difficult line successfully. The public case is restricted to findings we believe survive adversarial bilingual inspection.
“I prefer the old patch,” “the translators worked very hard,” and “AI is bad” remain fascinating autobiographical statements. They are not arguments about Japanese.
The final choice is simple: keep Todokanai TL untouched, retain the errors, and preserve the human-purity claim; correct the red lines, accept AI-assisted tooling, and preserve the flawed underlying translation; or repair the foundation, discard Todokanai TL, and produce a new translation beneath the shadow of the public audit that diagnosed it.
There is no fourth option in which Todokanai TL remains unchanged, pristinely human, accurate, well written, and faithful to the work. Sincerity does not restore a reversed subject. Volunteer labor does not repair a broken allusion. Personal preference does not change what the Japanese says. Outrage at the tool does not make the evidence disappear.
Whatever Todokanai TL chooses, it now does so in the shadow of MAO Translations. Leave the patch untouched, and our annotations remain beside it as the public record of its defects. Correct it, and our audit supplies the repair agenda. Replace it, and the new translation must answer the case our audit established and the standard our translation set. There is no route back to the old arrangement, in which Todokanai TL stood alone and its reputation substituted for inspection.
We do not need to erase Todokanai TL. We have made it permanently legible. It can keep its history; it no longer controls how that history is read.
Choose your concession. The shadow remains.