Interesting to see Anthropic now downplaying the new vulnerabilities that Mythos discovered:
> We reviewed a demonstration of this specific technique being used to identify a small number of previously known, minor vulnerabilities. These vulnerabilities all appear relatively simple, and we have found that other publicly-available models are able to discover them as well without requiring a bypass