BTW I was really impressed by the results of F5-TTS. The thing I liked best was the "Tagged" TTS, where you can specify a tag to use different tones of your own voice, like
{Angry}What have you done?
{Suprised}Me, I did nothing?
{Shouting}Who else do you think I'm talking to?
{Sad}Why are you always shouting at me?
I wonder if this would also work for "Character" tags, like {Susan}How was your day?
{Peter}I had a great day.
That would open great new ways of having audio books read by cloned voices - switching between characters with the same voice like often done by the real narrators