← Latest papers
📄 other

AI versus Expert Feedback on Ethical Decision-Making in Medical Students: A Pilot Randomized Controlled Feasibility Trial with Exploratory Educational Outcomes

This pilot randomized controlled trial demonstrates that AI-generated feedback is a feasible and acceptable alternative to expert panel feedback for medical ethics education, showing comparable or potentially superior preliminary outcomes despite baseline imbalances and the need for larger confirmatory studies.

Original authors: Tuğba İş Kara, Yavuz Selim Kıyak

Published 2026-07-25
📖 5 min read🧠 Deep dive

Original authors: Tuğba İş Kara, Yavuz Selim Kıyak

Original paper licensed under CC BY 4.0 (https://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer

Imagine the medical school classroom as a giant, high-stakes training ground where future doctors learn not just how to fix broken bones, but how to navigate the messy, gray areas of human life. It's a place where students must learn to balance rules, feelings, and tough choices—like deciding who gets a scarce resource or how to tell a family bad news. For decades, the best way to teach these "ethical muscles" has been to have a human expert, like a wise coach, review a student's choices and explain why a decision was good or bad. But here's the catch: finding enough wise coaches for every student is incredibly hard and takes up a massive amount of their time.

Enter Artificial Intelligence (AI), the digital brain that can chat, write, and solve puzzles at lightning speed. The big question researchers are asking is: Can this digital brain act as a substitute for the human coach? Specifically, can an AI generate feedback that helps students make better ethical choices just as well as a panel of real human experts? This isn't about replacing the human entirely, but rather seeing if a "human-reviewed AI" can be a scalable sidekick, offering the same guidance without burning out the faculty. If it works, it could mean every medical student gets a personal ethics tutor, no matter how large the class gets.


The Great Feedback Showdown: Human Coaches vs. AI Assistants

In a recent pilot study, researchers set up a fascinating experiment to see if a robot's advice could hold its own against a room full of human experts. They gathered 56 sixth-year medical students—basically interns who are about to graduate—and split them into two teams. Both teams played the same ethical video game called the "Concordance of Judgment Learning Tool" (CJLT). Think of this tool as a simulation where students face tricky scenarios, make a choice, and then get feedback on how their choice compared to what experts would do.

The twist? The two teams got their feedback from different sources. One group received notes written by a panel of 25 human experts. The other group received notes generated by an AI, but with a safety net: the AI's output was first checked and approved by human experts to make sure it wasn't hallucinating or being weird. The goal was to see if the AI group could learn just as well as the human group.

The Results: A Surprising (But Cautious) Lead for the AI

When the students took a test called the Objective Structured Video Examination (OSVE)—which is like a video quiz where they have to spot the right ethical move in a scene—the results were a bit of a rollercoaster.

First, there was a hiccup at the starting line. Before the training even began, the AI group was already scoring higher than the human group. It was like starting a race with one team already five meters ahead. The AI group's average starting score was 23.32, while the human group started at 19.36. This wasn't a tiny gap; it was a significant one.

After the training, both groups improved. The AI group ended up with an average score of 27.29, and the human group landed at 23.32. When the researchers used a statistical trick to adjust for that head start (the baseline difference), the AI group still looked like the winner. The adjusted math suggested the AI feedback might have given a boost of about 4.50 points compared to the human feedback.

However, the researchers are very careful not to pop the confetti just yet. When they looked at how much each student improved (the change from start to finish), the two groups were identical. Both groups improved by exactly 3.96 points. It was a dead heat. This suggests that while the AI group finished with higher scores, they might have just been better to begin with, rather than the AI teaching them something the humans couldn't.

What About the Other Tests?

The study also checked two other things: the Script Concordance Test (SCT), which measures how well students handle uncertainty, and a scale called BIAS, which looks at whether students have certain thinking shortcuts or prejudices. In these areas, the AI and the humans were neck-and-neck. There was no difference at all. The AI didn't beat the humans, but it also didn't lose. It performed just as well, or perhaps just as "average," as the human experts.

The Students' Verdict

The students who volunteered to give feedback three months later seemed to love the experience. Out of the 21 who responded, a massive 95.2% said they would recommend this training to other medical students. They found the scenarios realistic and felt it helped them think faster about ethical issues. One major hurdle they noted, however, was that in the real world, it's hard to challenge a boss or a senior doctor, even when you know they are ethically wrong.

The Bottom Line

So, did the AI win? The study suggests that AI-generated feedback is feasible and acceptable. It didn't fail; in fact, it didn't show any signs of being worse than human experts. The data even hints that it might be slightly better at helping students justify their decisions, but the researchers warn us to take that with a grain of salt. Because the AI group started with higher scores and the groups were small, we can't say for sure that the AI is the superior teacher.

Think of this study as a successful dress rehearsal. It proved that an AI coach, supervised by humans, can run the show without crashing the system. But to know if it's truly the MVP, we need a bigger stadium, more players, and a fairer starting line. For now, the AI is a promising sidekick, ready to help human experts teach the next generation of doctors how to do the right thing.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →