Title being updated from its current form to add that this is specifically voice and video would make it quite a bit more interesting. Using text to invoke an AI from Signal is not really interesting at all, that’s basically OpenClaw. Using voice and video to do so, however, is quite a bit more interesting