It is a very well known truism that it is harder to debug code than it is to write it. This is, in my experience, doubly true of the convoluted code generated by current frontier LLMs.
Allowing LLM coding at our org basically stole most of this year's progress from us as every PR made this way still has yet to be merged because the code quality just never reaches anywhere near our minimum requirement.
We are investigating ways to improve this (a style guide for agents etc) but the best step we've taken so far is just to ask people to stop using it and see what happens. (Code quality jumped up and PRs started getting merged, though the LLM ones are still languishing and probably will need rewriting from scratch before we can merge them)