Certainly big for China, but if it can't run CUDA then it's not going to help them catch up in the AI race
Even useful models like gemma3 27b are hitting 22 t/s on 4bit quants.
You aren't going to be reformatting gigabytes of PDFs or anything, but for a lot of common use cases, those speeds are fine.