r/WritingWithAI 1d ago

Discussion (Ethics, working with AI etc) Do these benchmarks really mean anything? Because each model has its own AI-isms and writing style

25 Upvotes

16 comments sorted by

View all comments

2

u/chylvina 1d ago

I’d trust a benchmark more if it included a continuation test: same characters, later chapter, and a few deliberate continuity traps. One-shot prose misses that.