Helping 'bad' people drew disapproval from some of 21 AI models
PhilosophyPsychology
THE AI ANGLE
Judging interpersonal morality and assigning reputationsResearchers tested 21 language models on social dilemmas and found conflicting judgments about whether to help people with bad reputations. Most models leaned toward a standard where cooperating is always good, though their evaluations also shifted with recipient gender and culture. These results show scholars that systems offering personal advice run on divergent moral frameworks that resist simple prompt corrections.
THE TEACHING ANGLE
Discussion can center on whether a community builds stronger cooperation by helping wrongdoers or by requiring members to shun them.Read the original at techxplore.com Generate teaching or study materials
More in Philosophy
- What do you do with an AI conscientious objector?CIO.com · September 29, 2026
- Misleading AI-generated summaries can distort human memoryPhys.org — Technology · September 29, 2026
- The Hidden Side of Student AI Use: Mental Health, Access and RiskEdTech Magazine — K-12 · September 29, 2026
- If AI makes the decision, who owns the consequence?CIO.com · September 28, 2026
- Podcast: The AI Revolution Fails Without Psychological Safety For Developers: A Conversation with Erin DoyleInfoQ · September 28, 2026