Skip to content

Why AI-Written Books Fail – and How We Made Quality Measurable

with Hannes Steiner & Thomas Neumayer

Thursday 24 September 2:50 PM – 3:20 PM Stage 3

About This Session

Anyone can generate 40,000 words with an LLM today. Almost none of it survives contact with a reader. At story.one, we generate complete non-fiction books – 17 chapters, printed and sold in real bookstores – and over the past 18 months we read and scored more than 1,000 of them, cover to cover. This talk is the engineering story of what we found. Why long-form generation fails in ways chatbot demos never show: phrase tics, structural repetition, fabricated quotes, visible seams between regenerated passages. Why the model is the smallest part of the problem. And how we built a release standard around a neurosymbolic idea: fast neural generation, slow symbolic verification – deterministic checks in code (quote verification against original sources, style-pattern detection, structural gates), blind evaluation with fresh context, and a nightly regression suite of real-world cases. You will leave with three transferable lessons for any long-form GenAI product – and you will see a book that was generated, verified, printed, and sold the same day.