You list aeneas as a dependency here: https://subaligner.readthedocs.io/en/latest/acknowledgement....
That makes me wonder how much of the work is being done by aeneas vs. your own model. If parts of the audio are in a low voice or noisy (which tends to cause aeneas to slip) will subaligner be able to fix that?