The Short Version
Generative AI models do produce factual inaccuracies, and this is a well-documented, persistent challenge confirmed by peer-reviewed research and major benchmarks. However, the word "consistently" overstates the problem. Error rates vary enormously — from below 1% on grounded summarization tasks to over 30% on open-domain reasoning — depending on the task, domain, model, and whether retrieval tools are used. Hallucination rates are also declining over time. The claim describes a real issue but frames it in a misleadingly uniform way.