Another thought I had: for performance reasons, it might be nice to have something more compact than a one-hot vector for each letter. Have you looked at determining sets of characters which have a similar impact on hyphenation, and encoding them together?
PS: do you have the extracted list of wiktionary hyphenations sitting in a text file somewhere that you could put up? I'm fixin' to quickly compare the accuracy to TeX's German hyphenation (once the 30+GiB TeXLive repository finishes downloading).
PPS: You could improve the display of code blocks in your site on desktop by adding
display: block;
max-width: 710px;
width: 80%;
margin-left: auto;
margin-right: auto;
to your `.post-content pre code` rule. Or maybe slightly indent it by reducing the max width a small amount below that of the body text.