People who don’t want to hear from Grok can just not use it. People who are concerned about what Grok might say to me can rest assured their concern is misplaced.
In this case, bias and unfairness can be a real concern since they will affect the products and the business. Its not about being 'preached to' but about avoiding pissed off customers and decrease of trust from them if our system generates hateful/biased text.
I want to emphasize that it's not about "what Grok might say to me". Of course, not. I've got friends who revel in dark humour and I part-take more than most. But if you've spent millions training an LLM, you're not just going to use it as a toy. It is going to solve some usecase. That usecase will probably affect real people. If your LLM is biased, or has little guardrails, those affects might not all be positive.
[1] https://research.ibm.com/blog/retrieval-augmented-generation...
But yes, you're right. We've seen how dangerous open systems can be. Hackers and scammers have shown us that we need guardrails against operating systems and the web, as you correctly point out. It's time for some legislation that locks the down so the unwashed masses don't have unfettered access to them, don't you agree?
Anyway I think it’s fine for these systems to be available as open source, I’m not suggesting they be withheld from the public. But when you offer it as a cloud service people associate its output with your brand and I think this could end up harming Twitter’s brand.
Microsoft did not like what they got and shut it down because it ended up being a 4chan troll.
a generic instruction-tuned LLM won't act like that.
Instruction tuning is on top of the base LLM and is often RLHF to train the base LLM to produce certain kinds of responses.
Much better.
[1] https://aclanthology.org/P19-1339.pdf [2] https://arxiv.org/pdf/1906.09208.pdf [3] https://developers.google.com/machine-learning/glossary/fair...
To be very clear, I want ALL queries (I presume you mean LLM prompts) to be available to everyone.
Could you also explain what do you mean by people like me? Indians? NLP researchers? People in their thirties? Expatriates?
OP mentioned nothing about fairness; that's orthogonal to objectivity, and you're projecting your worldview onto OP's.
The "harm" seems to be the public (media) outrage at some inappropriate content produced. Most companies cave immediately at any level of pressure, but I have a feeling Musk won't. He's the type to take it to court if needed.
PS: Yes I know CrowsPairs is a dataset with a bunch of flaws. My SO is working, in a team of 10+ linguists and researchers to develop a multi-lingual, generalized version of it which also addresses multiple problems with it. Unpublished work, for now.
[1] https://github.com/nyu-mll/crows-pairs/ [2] https://arxiv.org/abs/2004.09456 [3] https://arxiv.org/pdf/2204.09591.pdf