There's plenty of fine-tuning and RLHF involved too, that's mostly how "model alignment" works for example.
The system prompt exists merely as an extra precaution to reinforce the behaviors learned in RLHF, to explain some subtleties that would be otherwise hard to learn, and to fix little mistakes that remain after fine-tuning.
You can verify that this is true by using the model through the API, where you can set a custom system prompt. Even if your prompt is very short, most behaviors still remain pretty similar.
There's an interesting X thread from the researchers at Anthropic on why their prompt is the way it is at [1][2].
[1] https://twitter.com/AmandaAskell/status/1765207842993434880?...
[2] and for those without an X account, https://nitter.poast.org/AmandaAskell/status/176520784299343...
https://www.anthropic.com/research/constitutional-ai-harmles...
Then play whack-a-mole until you get what you want, enough of the time, temporarily.
From computer’s doing exactly what you state, with all the many challenges that creates
To is probabilistically solving for your intent, with all the many challenges that creates
Fair to say human beings probably need both to effectively communicate
Will be interesting to see if the current GenAI + ML + prompt engineering + code is sufficient
I can absolutely put into words what I want, but I cannot program it because of all the variables. When a computer can build the code for me based on my description... Holy cow.
That seems like a massive advantage.
I never thought my English degree would be so useful.
This is only half in jest by the way.
Scientists fold proteins, _hoping_ that they'll find the right sequence, based on all they currently know (best guess).
Without hope there is no need try; without trying there is no discovery.
When it says “please open iPhone to see the results” - half the time I think it’s capable of responding with something but Apple would rather it not.
I’ve always seen Siri’s limitations as a business decision by Apple rather than a technical feat that couldn’t be solved. (Although maybe it’s something that couldn’t be solved to Apple’s standards)
So there's direct evidence of Apple insiders thinking Siri was pretty great.
Of course we could assume Apple insiders realised Siri was an underwhelming product, even if there's no video evidence. Perhaps the product is evidence enough?