Don't forget that humans have a not insignificant error rate when copy/pasting or copy/typing data.

And it's possible to run each document through the LLM pipeline multiple times, using different models and/or prompts each time, to check for errors and inconsistencies. That will take more time and cost more, but it can reduce the error and hallucination rate significantly.