Will better AI models solve this?
Only partly. Larger models read existing code better, but the information they are missing (why the code is the way it is) was never recorded anywhere. That is a documentation problem, not a model problem.
From Prompt-to-app tools: what they do, where they stop, published 9 August 2026. That article is where the reasoning behind this answer is set out, including what it does not cover.