The core of the transcription process is mostly the same and uses the same libraries, but I've rebuilt the UI from scratch to be fully responsive, added support for Wayland/Hyprland (with adaptive window size), implemented lazy loading/filtering/group by date/editing/search/filtering etc to the history screen, implemented history storage/handling differently, added more control over the history feature, added support for custom sounds, improved UX around managing sounds, added a loading screen, added support for model download pause/resume/cancel/delete etc.
These might seem like details, but it all takes time. I started this project 3 weeks ago and this is just the beginning.
In my roadmap I've listed many ideas I have in mind and will be focusing on: https://docs.voice-ai.knowii.net/roadmap
I want to go in a different direction than Handy, and my customers (who are mainly interested in Knowledge Management) too.