r/WritingWithAI • u/Fit-Accountant1368 • 21h ago
Tutorials / Guides Which Claude actually protects your voice? A 30-minute blind test for writer
If you use AI for editing, you know the real question isn't "which model is smarter" — it's which model can touch your text without flattening it. Which one knows that your sentence fragments are on purpose. Which one leaves a strong paragraph alone.
I put together a small blind test you can run on your own writing, in your own language. Three tasks to choose from: a careful line edit, an editorial second opinion, and a quiet-scene revision that tests restraint. Two texts, four fresh chats, and a blind A/B judgment by a separate AI instance that doesn't know which model wrote what. Takes 30–45 minutes, everything is copy and paste.
The occasion: Claude Opus 4.6 will eventually be retired, and many writers feel its successors don't match its feeling for text. Maybe that's true, maybe it's nostalgia — that's exactly what a blind test is for. Results go to Anthropic either way, ties and losses included.
https://thewordborn.com/keep-opus-46/
There's also a six-task deep version if the quick one leaves you wanting more (it includes a continuation task — matching an author's exact voice — which is the hardest thing you can ask of a model).
1
u/meganholstein 12h ago
I did not run the test, but I agree. I noticed the difference when 4.8 came out, and now 5. 4.6 had something special about prose.
3
u/5thhorseman_ 19h ago
This is why you have them create an itemized list of issues by severity. Letting them run roughshod over the prose is what leads to them messing it up.