you cannot plug and play random models. they are all different trained on different data and rl for different capabilities.
Anthropic's best models are very good, maybe the best in their category. But, they have direct competition. You can, in fact, just switch to Codex or Gemini or GLM. It mostly is plug and play. I have a preference but I also have options.
well they dont tell you that do they? there is no way to tell what model can and cannot do unless you extensivevly test it yourself and pray for the best.
Do they lack "testing strategy" to test their own alignment?
Can you share the you testing strategies that are letting you plug and play models.