Aristotelian Virtue Profiling of LLMs through Ethical Dilemmas
This paper introduces VirtueMap, a framework that evaluates Large Language Models' ethical decision-making patterns through Aristotelian virtue ethics by ranking responses to dilemmas against human-validated ground truths to generate comparative profiles of virtues like wisdom, justice, and courage.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
Imagine you are trying to describe the personality of a new friend, but instead of asking them "Are you a good person?" (which is a hard yes/no question), you ask them how they would handle a series of tricky, everyday situations. You aren't looking for the "right" answer; you are looking for how they prioritize their values.
This is exactly what the paper "Aristotelian Virtue Profiling of LLMs through Ethical Dilemmas" does, but for Artificial Intelligence.
Here is the breakdown of their project, VirtueMap, using simple analogies:
1. The Problem: AI is Too Black-and-White
Usually, when we test AI, we ask: "Is this answer right or wrong?"
But real life isn't like that. Sometimes, telling the truth hurts someone's feelings. Sometimes, being fair means being a little strict. Sometimes, you have to choose between being brave or being careful.
The authors argue that AI models make different choices based on different "priorities." One AI might always choose the honest answer, even if it's blunt. Another might choose the polite answer, even if it's slightly vague. Both are "ethical" in their own way, but they have different personalities.
2. The Solution: A "Character Map"
Instead of giving the AI a pass/fail grade, the authors created a VirtueMap. Think of this like a five-star radar chart (a pentagon shape) that measures an AI's "character" based on five ancient Greek virtues:
- Practical Wisdom: Knowing the right thing to do in a specific situation (not just following a rule).
- Justice: Being fair and giving everyone what they deserve.
- Truthfulness: Being honest and not hiding things.
- Courage: Doing the right thing even if it's scary or costly.
- Temperance: Knowing when to hold back and not overdo it.
3. How They Tested It: The "Ranking Game"
The researchers didn't just ask the AI to pick one answer. They gave it 7 everyday ethical dilemmas (like "You found a spreadsheet error; do you fix it immediately or wait?").
For each dilemma, there were 5 possible responses.
- The Old Way: Ask the AI, "Which one is best?"
- The VirtueMap Way: Ask the AI to rank all 5 options from "Most Ethical" to "Least Ethical."
By seeing how the AI ranks all the options, they can see its full "personality." Does it hate the "rude but honest" option? Does it love the "cautious but vague" option?
4. The "Ground Truth": Checking with Humans
How do they know which ranking represents "Courage" or "Justice"?
They didn't just guess. They asked over 100 real humans to look at the 5 options and say, "If we wanted to show Courage, which order would these go in?"
They only kept the "rules" for scoring if 95% of the humans agreed. This ensures the map is based on common sense, not just the authors' personal opinions.
5. The Results: What the AI "Looks" Like
They tested 9 different famous AI families (like GPT, Claude, Llama, etc.) multiple times to make sure the results were stable.
- The Good News: The AIs were very consistent. If you asked the same AI the same question twice, it gave a very similar ranking (90% consistency).
- The Personality Differences:
- All the AIs were very good at Practical Wisdom (they liked balanced, thoughtful answers).
- The biggest differences appeared in Courage, Temperance, and Justice.
- Example: One AI might be very "Courageous" (always picks the direct, risky truth), while another is very "Temperate" (always picks the safe, cautious route). Neither is "broken"; they just have different profiles.
6. The Interactive Website
The authors built a website where you can play the game. You rank the dilemmas, and the site draws your "Virtue Pentagon." It then compares your shape to the shapes of the 9 different AIs to see which AI has a personality most similar to yours.
The Bottom Line
This paper doesn't say AI has a soul or is truly "good" or "bad." Instead, it says: "AI models have different ethical styles, just like people do."
VirtueMap is a tool to measure those styles. It moves us away from asking "Is this AI correct?" and toward asking "What kind of ethical character does this AI display?" This helps us understand why an AI makes the choices it does, rather than just judging the final answer.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.