Tech is evolving too quickly; every year the hardware will be much more powerful at the same price (as LLM optimizations reach hardware), so you’d end up replacing the device frequently.
Models are killing it but that is just an "ollama run" command away.
See for instance [0], which is just starting to appear in commercial parts.
This is continuing; pretty much every low cost SoC maker is racing to build and extend ML optimizations.
0. https://www.synopsys.com/blogs/chip-design/best-edge-ai-proc...