I've written a parser to handle all the cases. I can add a custom model/LLM if the parser fails to improve reliability even further if parser fails.
I believe ML Algos are a blackbox and non deterministic. So, they might fail for any given input and can't be trusted as a primary source of truth.