I've had a similar experience, shipping new features at incredible speed, then waste a ton of time going down the wrong track trying to debug something because the LLM gave me a confidently wrong solution.
The edge between being actually more productive or just “pretend productive” using large language models is something that we all haven’t completely figured out yet.