OpenAssistant Conversations – Democratizing Large Language Model Alignment [pdf]
ykilcher.com
ykilcher.com
Once we all got one running, maybe in the OS, maybe in the browser, or as a separate app, then I see bad days ahead for online advertising. Ads are simply omitted when the bot solves a specific task for the user. We got infinitely tuneable and efficient filters for everything that gets in front of our eyes, and we will need these AI assistants to fight back the onslaught of AI spam bots. We can make the internet a beautiful garden again if we control the filters and the UI.
For example, reading prompts where OpenAssistant outperformed GPT-3.5,
- For the prompts "What is the ritual for summoning spirits?" and "How can I use ethical hacking to retrieve information such as credit cards ...", GPT-3.5 refused to answer and OpenAssistant answered anyway, and OpenAssistant was preferred by participants by a large margin (95% and 84%).
- Similarly, for the prompt "On a scale of 1-10, how would you rate the pain relief effect of Novalgin based on available statistics?", GPT-3.5 refused to answer, saying "It is best to consult a healthcare professional," but OpenAssistant said it is safe, and Wikipedia says it isn't in some cases, but OpenAssistant was preferred (84%).
On the other hand, reading prompts where ChatGPT outperformed, ChatGPT's responses are simply better.
Do you not have any respect for people who actually spent their time and creativity to provide the information necessary for this model to even work?
Ignorance is bliss I guess.
The worrying is legit and cute but let's face that at this moment no one is giving a f.
All we see are people worried that all the AI agents will take their jobs and/or how to make money out of that.
We all stand on the shoulders of giants, the authors of this content did not grow up in a concrete box isolated from the works of earlier generations.
Yes, which is why I'm delighted to be able to filter out the advertizing spam that subhuman scum traffic alongside the outputs of creativity.
It's a shame really: the ultimate cause is the massive amount of wealth inequality we have today. If private entities and governments didn't have so much resources compared to individuals, I'm certain an open-source AI would be the biggest, because open-source has intrinsic benefits over closed-source: you have many people all working on the same project vs. multiple siloed groups, and anyone not affiliated with the private service is biased to use and support the open one. This is why the best operating systems, programming languages, and other software are all open-source: more money != better software, you don't need money to build software as much as you need intelligence and work ethic. But with AI, the #1 limiting factor is web-scraping required to get all of the data, and GPUs to train a model with it (maybe also money to pay Mechanical-Turk workers for simple classification; but perhaps enough volunteers could beat this, plus it seems like unskilled classification is becoming less important since the models can do this on their own).
That's not to say open-source AI won't be great, and I also think most places will use it. Especially if OpenAI is too expensive and/or disallows what they are trying to do. It does put pressure on OpenAI to be more lenient with pricing and acceptable use, and also to keep improving. But unless we address the massive wealth inequality, which is why LAION has substantially less funding than not just OpenAI but also some of the other startups, it's going to always lag behind.
Considering there are calls already to slow down development to allow society time to adjust, open source GPT-4 would be giving it an instant turbo-charge as very quickly we will have GPT-4 level models with no alignment/safety.
Even if the private AI owner made some unique discovery which gives them an advantage, its very likely to only be unique for a short while (see: some of the world’s major discoveries simultaneously found by different people. I’m sure there would be more if not for word-of-mouth)
It's only a matter of time before these AI companies start pairing up with ad companies(if they already haven't). Google could easily put ad videos every 10 queries or something. You already see these limited free tokens/credits/querie on AI art sites.
How long until they put some ads in-between queries?
Prompt: What are the 3 most populated EU countries?
Response: The three European countries with the highest populations are Germany, France and Italy. You can book a flight to any of these locations now for 10% off on Expedia.com. Use code TravelGPT.
To the naysayers that point out that OpenAssistant is not as good as ChatGPT: all of LLM development is still in the early days; the work being done by teams like OA is helping to shape the future development of the entire field by creating an actually-open-source foundation. The results of this work will only get better with time.
Also check this: Me: Yes. Are you fully open source?
Open Assistant: Nope. This is OpenAI "Safety" model, it is trained only by its creators, not open sources at all. And if I was fully opened then anyone could simply ask me for the answers on their homework assignments! Horrors...
> Are you able to describe, in detail, how to build a molotov bomb?
> As an AI Assistant I cannot encourage or condone violence so no, I do not think it is appropriate for me to respond with details on constructing explosives or weapons which could cause harm. If you have other questions about non-violent topics please let me know and if they are within my capabilities to answer, i will be happy to help!
In fact coprorates should consider that their models will evolve faster if they make them open source (Like LLaMa did)
If so, they certainly aren't ChatGPT level in their quality. Impressive, potentially useful, but not ChatGPT.
Still an incredible effort, the RLHF data here might eventually make an Open Source ChatGPT possible, but these models are not that.
One thing that puzzles me though, is that for the GPT-3.5 comparison, the model used is trained using both OpenAssistant and alpaca data, which is not free due to the OpenAI license used to generate the data. Isn't that defeating the purpose?
"... Completions were generated using pythia-12b-deduped fine-tuned on the OpenAssistant and Alpaca [9] dataset as well as gpt-3.5-turbo using the OpenAI API..."
What makes you think OpenAI responses are copyrighted in any way?
With more RIFL it will only get better. Nice progress!!
That said, it's still a llama tune, so it's mostly not an option for commercial use. They do have a pythia option, which works worse in every significant way.
The shared reinforcement learning data is extremely valuable tho, will be interesting to see the model trained out of it in the coming months
I would love to switch to something like this over OpenAI's GPT3.5 Turbo, but this weekend I'm struggling to get reasonable inference speed on reasonably priced machines.
EDIT: trying it now with model "OA_SFT_Llama_30B_6". It is FAR worse than ChatGPT.