This is the problem with cloud models, you build a "predictable" workflow then they remove it with a new and improved one that is less deterministic and often costs more. If you use a local model discontinuation is no longer a thing to worry about.
It's not even the new one being "less deterministic and often costs more".
Application prompts overfit the model they are using to get the output they want. Switch model, the prompt no longer produces the output you expect and you may not be able to get the output you need anymore.
If a model is critical for your application/business you want the ability to run it yourself otherwise you are stuck when it becomes unavailable.