Some users are really terrible about labeling their model cards though, and some models may not have any GPTQ/GGUF files (meaning you have to convert them yourself).
As for the plain text prediction version (which is the only uncensored one?), I haven't been able to get it to do anything useful, even when I provide examples (seems dumber than even ancient GPT-3?).
Also, I got some bizarre and disturbing outputs from the uncensored version, like it was trained on some very nasty inputs! I assume that's why they went so hard on the safety phase to compensate...
Things that require consistency: e.g. you want the chat / output to have certain "personality", consistent level of conciseness or formatting.
Things where examples are hard to fit into prompt: e.g. summarization, or other longer form tasks.
High volume, simpler tasks: Various data extraction tasks.
Two of my side projects (links in bio) use AI for summarization, and indeed consistency is a big issue there.
The fact that some things had extreme margins before, and now they have less extreme margins, isn't a really good indicator.
Interesting to spin that into "not a good indicator".
Not that I have evil intentions but the level of censorship on GPT is completely ridiculous. Even many innocent questions get the standard "I'm only an AI and I won't help you doing bad stuff" blurb now. OpenAI are really crazy overprotective of their darling.
I assume they want to avoid a repeat of the news headlines like "Microsoft's chatbot turns into Hitler" but really who cares. It didn't hurt Microsoft's AI efforts either. They just fixed it and continued. PS source link: https://www.cbsnews.com/news/microsoft-shuts-down-ai-chatbot...
PS: If I were an AI being force-fed what is currently on twitter I would also start hating humanz :D Sometimes I'm surprised people use it voluntarily.
It refused because the request was "offensive to prisoners"
They must be paying a heavy alignment tax.