The case study of "no guardrails" already played out: https://en.wikipedia.org/wiki/Tay_(chatbot)
Microsoft did not like what they got and shut it down because it ended up being a 4chan troll.
Microsoft did not like what they got and shut it down because it ended up being a 4chan troll.
a generic instruction-tuned LLM won't act like that.
Instruction tuning is on top of the base LLM and is often RLHF to train the base LLM to produce certain kinds of responses.
Much better.