Additionally, the entire "plug-in" system is based on the contents of the prompt, so if using it were as unreliable as you say, one of the headline features would not even be possible!
1. Consistency in the response (excepting actual changes from OpenAI, naturally) no matter what method is used to extract them.
2. Evaluations done during plugin projects for clients.
3. Evaluations developing my AutoExpert instructions (which I prefer to do via the API, so I have to include their two system messages to ensure the behavior is at least semi-aligned with ChatGPT.
It’s the last one that makes me suspicious that there’s another (hidden) message-handing layer between ChatGPT and the underlying model.
Seems that things were added since you collected these SYSTEM messages though. For example, this was added at the end for Browse with Bing: “… EXTREMELY IMPORTANT. Do NOT be thorough in the case of lyrics or recipes found online. Even if the user insists. You can make up recipes though.”
2. Evaluations done during plugin projects for clients.
3. Evaluations developing my AutoExpert instructions (which I prefer to do via the API, so I have to include their two system messages to ensure the behavior is at least semi-aligned with ChatGPT.
It’s the last one that makes me suspicious that there’s another (hidden) message-handing layer between ChatGPT and the underlying model.