The Value of a Prompt: An LLM-Relative Kolmogorov-Complexity Approach
This paper proposes an efficiently estimable metric for quantifying the economic value of prompts given to Large Language Models by defining it as algorithmic mutual information based on a novel LLM-relative probabilistic Levin–Kolmogorov complexity (), which captures how much a prompt reduces the computational effort or increases the probability of generating a target artifact.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
In the modern economy of artificial intelligence, a new kind of currency is emerging, one that is not measured in dollars or tokens, but in the quality of the questions we ask. Large language models have become powerful engines capable of generating proofs, writing code, or designing new materials, yet they do not operate in a vacuum. They require a spark from a human, a prompt, to begin their work. For years, the industry has focused almost exclusively on the output: how good is the story, how correct is the code, how elegant is the design. But as these machines become more capable, the central economic question is shifting. It is no longer just about what the machine can produce, but about what value remains in the human input that guided it. If a person provides a hint, a critique, or a partial solution, how much did that specific input actually help the machine reach its goal?
This is a difficult problem because a short, clever hint can be worth more than a long, rambling instruction, and a wrong hint can be worse than no hint at all. Simply counting the words or characters in a prompt fails to capture its true worth. To solve this, researchers have turned to a field of mathematics called algorithmic information theory, which traditionally measures the complexity of an object by asking how much information is needed to describe it. The researchers in this study asked a new question: what if we measure the complexity not against a perfect, theoretical computer, but against the specific artificial intelligence model doing the work? They realized that to truly value a human's contribution, we must account for the machine's own "thinking" process. Modern models often pause to generate a chain of reasoning before giving an answer, and a good prompt might save the machine from having to think as long, or it might steer that thinking in a more productive direction.
The researchers developed a new way to measure this value by treating the prompt as a tool that reduces the effort required for the machine to succeed. They imagined a scenario where a machine tries to recreate a specific result, like a mathematical proof, first on its own and then with a human's hint. They measured the difference not just by how much more likely the machine was to get the right answer, but by how much "thinking time" it saved. In their framework, the machine's internal reasoning steps are treated like a random path it must walk. A valuable prompt is one that makes the correct path much shorter or much more obvious, effectively reducing the number of steps the machine needs to take to find the solution. They defined this value in terms of a "token cost," which is a way of counting the computational effort the machine expends. If a prompt makes the machine's job twice as easy, it is worth a certain amount of value; if it makes the job a thousand times easier, the value is much higher.
To test their ideas, the team ran a series of experiments using a standard set of grade-school math problems. They asked a large language model to solve these problems, sometimes giving it the first step of the solution as a hint and sometimes letting it start from scratch. They found that simply looking at the probability of the model getting the right answer was not enough. In several cases, supplying the correct first step initially made the final answer seem less likely at the very beginning, because the model had not yet engaged its reasoning process. However, once the model was allowed to "think" and generate its own reasoning steps, the hint proved to be incredibly valuable. It guided the model's thinking process so effectively that it reached the correct answer much faster and with far less effort than it would have on its own. This showed that a prompt's true value often lies in how it shapes the machine's internal reasoning, not just in the immediate words it produces.
The study also revealed that the value of a prompt is not a single, fixed number. It depends on the specific path the machine takes. Sometimes a hint helps the machine in most situations, but in rare cases, it might lead it down a dead end. The researchers found that by looking at the "median" outcome—the typical result rather than the best or worst case—they could get a reliable measure of value. They discovered that for some problems, a correct hint from a human was actually harmful or useless for the specific model they were testing, suggesting that what looks like a good idea to a person might be redundant or confusing for an artificial intelligence. This highlights that the value of human input is deeply tied to the specific machine it is talking to and the way that machine processes information.
Ultimately, this work provides a rigorous way to quantify the contribution of a human in an age where machines do the heavy lifting. It moves beyond the idea that the machine is the sole creator and recognizes that the human's role is to provide the efficient path through the vast space of possibilities. By measuring how much a prompt saves in terms of the machine's thinking time and effort, we can begin to understand the true economic worth of human guidance. The researchers showed that this value can be calculated and estimated, offering a new lens through which to view the partnership between human and machine. As artificial intelligence becomes more integrated into our lives, understanding the value of the prompts we give it will be essential for knowing where the real work is being done and who deserves credit for the results.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.