The ‘safe’ tuning of the models is becoming a nuisance. As indicated in the paper, the agents are overly cooperative and pleasant due to the LLM’s training.
Pity they can’t get access to an untuned LLM. This isn’t the first example I’ve read it where research is being hampered by the PC nonsense and related filters crammed into the model.
Pity they can’t get access to an untuned LLM. This isn’t the first example I’ve read it where research is being hampered by the PC nonsense and related filters crammed into the model.