I had the same concern. However, the structure of the output was surprisingly stable. We rejected badly formatted responses: https://github.com/outerbounds/hacker-news-sentiment/blob/ma...
The semantics of the topics/tags could be improved for sure with a more detailed prompt