Steering LLMs Responses Towards Moral Foundations on the Norwegian MFQ-30
Researchers explored the feasibility of using the Norwegian Moral Foundations Questionnaire (MFQ-30) to steer large language models (LLMs) towards moral foundations. They administered the questionnaire to six LLMs and compared their responses to a sample of human respondents. The results showed that a neutral persona prompt can bring the LLMs' moral foundation profiles closer to the human mean, with some models showing significant improvement. This research has implications for the development of more human-like and morally aligned AI agents.
Save an API key to vote.