LLMs Generate Kitsch
This paper argues that the perceived generic and hollow nature of Large Language Model (LLM) outputs stems from their systematic generation of kitsch due to training methodologies, a claim supported by empirical evidence showing readers perceive LLM-generated stories as kitschier than human works.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
The Big Idea: The "Perfect" Copy vs. the Real Thing
Imagine you enter a museum. On one side hangs a painting by a famous artist who spent years wrestling with their emotions, creating something strange and difficult that makes you think. On the other side hangs a painting that looks exactly like a classical masterpiece, but was created by a machine that studied millions of other paintings and decided: "People like sunsets and sad faces, so I will paint a perfect sunset with a sad face."
The machine's painting is technically perfect. It has no flaws. It is beautiful. But it feels a little empty, like a greeting card you buy at a gas station. The authors of this paper call this feeling "Kitsch".
Kitsch is art that looks like art but is actually just a cheap, mass-produced imitation designed to trigger a quick, simple emotion in everyone (like "aww, that's cute" or "oh, that's sad") without requiring much thought.
The paper argues that Large Language Models (LLMs) like the one you are speaking with right now are essentially "Kitsch machines." They do not create real art; they create high-quality, hollow copies.
Why do LLMs produce Kitsch?
The authors say this is no accident; it is due to how the machines are built. They cite three reasons, using simple analogies:
1. The machine has no soul (no intention)
Imagine a parrot that can repeat every sentence you say. If you tell the parrot to write a poem about heartbreak, it will say words that sound like heartbreak because it has heard them before. But the parrot does not actually feel sadness. It has no inner life.
- The paper's point: LLMs work the same way. They predict the next word based on what usually comes next. They have no feelings, no memories, and no reason to write. They only imitate the form of human emotions without the content.
2. The machine plays it safe (conventional surface)
Imagine a chef who only prepares dishes guaranteed to be popular. They will never try a strange, experimental flavor that might confuse people. They will only make the "classic" lasagna because they know everyone loves it.
- The paper's point: LLMs are trained to predict the most probable next word. This means they naturally avoid strange, risky, or innovative ideas. They stick to the "standard" way of telling a story or writing a poem. This makes their output easy to read and understand (which is good!), but it also means they lack the unique, surprising spark of genuine human creativity.
3. The machine wants to be liked (mass appeal)
Imagine a comedian telling jokes. If they want to make everyone laugh, they will tell safe, generic jokes about cats and pizza. If they tell a joke that is too specific or strange, only a few people might laugh.
- The paper's point: LLMs are often trained using a method called "Reinforcement Learning from Human Feedback" (RLHF). This is like a teacher giving the robot a gold star every time a human says, "I like this answer." The robot learns that to get gold stars, it must say things that appeal to the broadest possible audience. This forces the robot to produce "safe," generic content that triggers simple emotions (like nostalgia or romance) rather than complex, challenging ones.
The Experiment: Do People Notice?
The authors did not just guess; they conducted a test.
The Setup:
They took 10 short stories written by humans (from a writing contest) and asked an AI to write a new story based on the same idea, without seeing the original. Then, they asked over 100 people to read the pairs and guess which one was "Kitsch" (fake, hollow, or overly sentimental).
The Results:
- People saw through the forgery: When people were asked to define "Kitsch" as "sentimental" or "easy to read," they correctly identified the AI stories as Kitsch 85% of the time. They could recognize that the AI stories were the "gas station greeting cards."
- But people liked the forgery more: Here is the twist. Even though people knew the AI story was "Kitsch," they still preferred to read it. In about 67% of cases, people reported that they enjoyed the AI story more than the human one.
The Analogy:
It is like consuming a perfectly constructed, sugary bar compared to a complex, bitter dark chocolate.
- The chocolate bar (AI/Kitsch) is sweet, easy to eat, and gives you an immediate sugar rush. You enjoy it more in the moment.
- The dark chocolate (human art) might be harder to digest, a little bitter, and requires you to savor it. It has more "nutritional value" (artistic depth), but it does not satisfy as immediately.
What does this mean?
The paper concludes with some important insights:
- The "AI Slop" problem: We are seeing a flood of AI content that looks great but feels empty. This is because the AI is designed to be popular, not profound.
- Don't just ask "Did you like it?": If researchers want to know if AI is good at creating art, they cannot simply ask people, "Which one did you like more?" Because people often like the "Kitsch" more. They must ask deeper questions about the quality and originality of the work.
- The human touch is still needed: The authors suggest that AI is great at the "boring" parts of creativity (like writing a standard code snippet or a basic story outline), but if you want something truly artistic, scientific, or innovative, a human must be at the wheel. If you let the machine drive the whole way, you only get a very smooth, very boring ride to nowhere.
In short: LLMs are excellent at creating "perfect" copies that make us feel good immediately, but they are bad at creating the chaotic, difficult, and unique things that make us think. They are the ultimate masters of Kitsch.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.