What is wrong with this? Tests involve a lot of hardcoding and mocking. I see this as an excellent use case for AI.
I deeply hate "regression tests" that turn red when the implementation changes, so you regenerate the tests to match the new implementation and maybe glance at the diff, but the diff is thousands of lines long so really it's not telling you anything other than "something changed".