Seems the trend with most LLM tech is once the users are there the model is quantized or downgraded down to the bare minimum level of usefulness either silently or through new versions that are just not better in real use.
That is to say, the release version of ChatGPT GPT4o seemed much better than the version today. This does not apply to the API.