Evaluating the relationship between regularity and learnability in recursive numeral systems using Reinforcement Learning
Using reinforcement learning, this study demonstrates that the high regularity of human recursive number systems facilitates their learnability from limited data, suggesting that this structural feature is a primary driver of their cross-linguistic spread.
Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer
Imagine you are trying to teach a robot how to count. You have two options for the "language" of numbers you give it:
- The "regular" language: Like English or Mandarin. Once the robot learns that "ten" plus "one" is "eleven" and "ten" plus "two" is "twelve," it can easily guess the rest. It is like a cookbook where the instructions are consistent: "Add one cup of flour, then add one egg."
- The "irregular" language: A made-up system where the rules change randomly. Maybe "ten plus one" is "eleven," but "ten plus two" is "banana" and "ten plus three" is "purple." There is no pattern to reuse; every number is a brand-new, confusing secret key.
This work poses a simple question: Is the reason humans use regular number systems (like English) that they are easier to learn?
To find out, the researchers did not ask humans; they built a digital "student" (a reinforcement-learning agent) and gave it thousands of different number systems to learn. They measured how quickly the robot could master the system.
Here is what they discovered, broken down into simple concepts:
1. The "Test" is Most Important
The researchers found that how they tested the robot changed the results.
Scenario A (The "Reality Test"): They trained the robot mainly with small numbers (1–10), since humans use these most frequently. Then they tested it on all numbers up to 99, including the rare, large ones.
- The Result: The robot learned the regular systems much faster. Because the rules were consistent, the robot could "generalize" (apply the logic) what it learned about small numbers to the large, rare numbers. The irregular systems were a nightmare; the robot had to memorize every single number from scratch.
- The Analogy: Think of regular systems like a Lego set. Once you know how to connect two bricks, you can build a tower, a car, or a castle. Irregular systems are like a box of random, unique puzzle pieces, where each piece fits only in one specific spot. If you want to build a large structure, the Lego set is much easier to learn.
Scenario B (The "Frequency Test"): They trained and tested the robot under the same skewed distribution (mainly small numbers, almost never large ones).
- The Result: In this specific case, it did not matter to the robot whether the system was regular or not. Since it was rarely tested on large numbers, it did not need to generalize. It simply memorized the small numbers it saw often.
- The Insight: Regularity helps you learn only if you later need to be able to handle new or rare situations.
2. The "Too Strange" Zone
The researchers also examined systems that were extremely chaotic and irregular.
- The Result: With these "super strange" systems, regularity played almost no role. Instead, the main problem was length. If the words for numbers were too long and complicated to pronounce, the robot had difficulty learning them, regardless of whether there was a pattern or not.
- The Analogy: If you try to memorize a list of random words, it does not matter whether they rhyme (regular) or not. If the words are simply incredibly long and hard to pronounce, you will forget them anyway. But if the words are short, a rhyming scheme (regularity) makes them much easier to remember.
3. The "Local" Pattern
The researchers found that a system sometimes looks chaotic overall but has small pockets of order.
- The Result: The robot learned best when the rules were locally consistent. For example, if the rules for numbers 30–39 were consistent, even if the rules for 40–49 were different, the robot could still learn it well.
- The Analogy: Imagine a city where every street in the "North District" follows a grid, but the "South District" is a labyrinth. You can still navigate the city easily if you simply stay in the North District. You do not need the entire city to be a grid to get around; you only need some local consistency.
The Big Conclusion
The work concludes that human number systems are likely not regular just because they are efficient to speak, but because regularity makes them easier to learn and pass on to the next generation.
If a language's number system is too chaotic, it is difficult for a child (or a robot) to figure out the rules for large numbers without being explicitly taught every single one. Regularity acts like a safety net, allowing learners to take a few examples and guess the rest correctly. This "learning pressure" is probably why human cultures worldwide naturally tend to use regular, pattern-based number systems.
Drowning in papers in your field?
Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.