Sorry, I missed your reply!
Local LLMs previously weren't good enough, but I've recently done another set of benchmarks (https://nuenki.app/blog/llm_translation_comparison) and llama 3.3 70b is getting there with some languages. Currently I'm in the process of integrating it via Groq, but it could plausibly be done with local LLMs.
Above a certain scale I'll start self hosting, but I'm nowhere near justifying those fixed costs yet.
Would you prefer to run ollama locally, or just know that the translation server is running its own models rather than forwarding onto cloud providers?