Paste the prompt and fill the five placeholders. verified_facts is the one that decides whether the output is worth anything: with it, the editor can swap "significantly faster" for a real figure; without it, every soft claim lands in the "Needs a fact" list and the edit is structural only. That is the intended behaviour, not a failure, but it means an empty verified_facts gives you a tighter draft rather than a more credible one.
Check the output in one pass: read the change log's right-hand column and confirm every specific in it appears in your input. This is where fabrication shows up. The usual shapes are a number that arrived from nowhere ("up to 40 percent faster"), a named study that does not exist, and a first-person anecdote about a client or a warehouse the model has never seen. Anything in the edited draft that traces to neither the original nor verified_facts is invented, and the change log is the fastest place to catch it because the model has to write down what it put there.
Work in chunks of roughly 800 to 1,200 words. On a full 3,000-word article the model starts summarising instead of editing, and the tell is the change log: it covers the first third in detail and then thins out. Rerun the back half on its own rather than asking for a longer log.
The other failure to watch for is over-correction. Told to break a symmetric rhythm, a model will sometimes produce a page of short declarative sentences, which is just a different uniform rhythm. If every sentence in the output is under twelve words, paste the edited draft back in with voice_sample filled and ask only for sentence-length variety.