← Latest papers
💬 NLP

compar:IA: The French Government's LLM arena to collect French-language human prompts and preference data

The paper introduces compar:IA, an open-source French government platform that collects large-scale, French-language human preference data through blind pairwise comparisons to address the scarcity of non-English datasets for training and evaluating Large Language Models.

Original authors: Lucie Termignon, Simonas Zilinskas, Hadrien Pélissier, Aurélien Barrot, Nicolas Chesnais, Elie Gavoty

Published 2026-02-09
📖 4 min read☕ Coffee break read

Original authors: Lucie Termignon, Simonas Zilinskas, Hadrien Pélissier, Aurélien Barrot, Nicolas Chesnais, Elie Gavoty

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer

Imagine the world of Artificial Intelligence (AI) as a massive library where the books are written almost entirely in English. If you try to read a book in French, Spanish, or another language, the story might feel clunky, the characters might act strangely, or the advice might not make sense culturally. This happens because the AI "students" spent most of their time studying English books and rarely practiced with others.

To fix this, the French government built a special playground called compar:IA. Think of it as a giant, public blind taste-testing contest for AI models, but instead of ice cream, they are tasting answers to questions.

Here is how it works, broken down into simple parts:

1. The Blind Taste Test (How it Works)

Imagine you walk into a room and see two mystery chefs (AI models) cooking up an answer to your question. You don't know who they are yet.

  • Step 1: You ask a question (like, "How do I bake a croissant?" or "Tell me a joke").
  • Step 2: Both chefs write a response. You read them side-by-side.
  • Step 3: You pick the one you like better.
  • Step 4: Only after you vote do the chefs reveal their names.

This "blind" method is crucial. It stops you from picking a chef just because you recognize their name or brand. It ensures the vote is based purely on the quality of the answer.

2. Why Build This? (The Problem)

Most AI training data is like a diet of only English food. Because of this, AI struggles with French culture, slang, and safety rules. To teach an AI to be good at French, you need thousands of French people to say, "I liked this answer, but I hated that one."

Before compar:IA, this data was either hidden inside big tech companies or simply didn't exist for French speakers. compar:IA is a public kitchen where anyone can help cook up this data, and then the recipe (the data) is shared freely with everyone.

3. The Harvest (What They Collected)

Since opening in late 2024, this digital playground has been busy. As of early 2026, they have gathered:

  • 600,000+ questions asked by real people.
  • 250,000+ votes deciding which AI answer was better.
  • 89% of the data is in French, which is a huge deal for a world dominated by English data.

They didn't just collect the votes; they also asked users to react to specific parts of the answers, creating a rich map of what French speakers actually care about.

4. Who Uses It and Why?

  • For Regular People: It's a free way to chat with many different AI models (some famous, some new and open-source) to see how they think. It also shows them how much energy (electricity) the AI uses to answer, encouraging them to think about the environment.
  • For Scientists and Companies: They can download the "recipe book" (the data) to train their own AI models to speak French better.
  • For Schools: Teachers use it to show students how AI works, turning the platform into a classroom tool where students debate which AI answer is the fairest or most accurate.

5. The Rules of the Game (Privacy & Safety)

The creators were very careful about privacy.

  • No Sign-ups: You don't need to give your name or email to play. This lowers the barrier to entry but means they have to be extra careful later.
  • The Filter: Before any data is shared with the public, a smart computer scans every conversation. If it spots a real person's name, address, or private info, it throws the whole conversation in the trash. It's better to lose a few data points than to accidentally leak someone's secrets.
  • Open Source: The data is released under a French government license, meaning anyone can use it for research or building new tools, as long as they respect the rules.

6. The Scoreboard

The platform also keeps a leaderboard (a ranking list). It's like a sports scoreboard that shows which AI models French people prefer the most. However, the authors warn that this isn't a perfect test of "intelligence." It just shows what people liked in a casual setting. It's more of a popularity contest than a final exam.

7. The Future

The French team sees this as a prototype. They are already working on expanding the "taste test" to other languages (starting with Danish) so that other countries can build their own public data gardens.

In a nutshell: compar:IA is a public service that crowdsources French language preferences to teach AI how to speak, think, and behave like a human French speaker, while keeping the process open, private, and free for everyone to use.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →