The issue is generally not accuracy in my experience (humans are even lazier about this). The issue is that LLMs do not understand what is interesting and what is procedural.