Mine too! I was not serious. I was just answering your napkin joke. But when I reread now, it is not that obvious. I guess with text you can pass only so much of emotions.
Considering this is a 14MB model running almost in a microcontroller, I am fine with such ‘ambigiuous’ queries cannot be handled, as long as the model confidence score is accurate. By the way, I did not test this model thoroughly. I am speculating on the potential of a small model like this. I don’t know if this one is good enough or not.
Tested your example, the confidence is 0. In smart home context, I can think of an application where the low confidence answers can be forwarded to cloud, whereas the vast majority generic queries solved locally, if the confidence is reliable enough. The response is quite fast by the way.
Engineering is something you calculate and get expected results with some acuracy. E.g building a bridge without trials and errors (hopefully). I would prefer prompt “tailoring” which is more like you start with something general and try to fit it to your desired output with many many trials and errors.