I've been wanting a local LLM appliance.
Models are killing it but that is just an "ollama run" command away.
See for instance [0], which is just starting to appear in commercial parts.
This is continuing; pretty much every low cost SoC maker is racing to build and extend ML optimizations.
0. https://www.synopsys.com/blogs/chip-design/best-edge-ai-proc...