← Latest papers
💬 NLP

Tendem: A Hybrid AI+Human Platform

Tendem is a hybrid AI-human platform that leverages automated agents for structured tasks and expert human intervention for verification, achieving superior quality and speed compared to AI-only or human-only workflows while maintaining comparable operational costs.

Original authors: Konstantin Chernyshev, Ekaterina Artemova, Viacheslav Zhukov, Maksim Nerush, Mariia Fedorova, Iryna Repik, Olga Shapovalova, Aleksey Sukhorosov, Vladimir Dobrovolskii, Natalia Mikhailova, Sergei Tilga

Published 2026-02-03
📖 5 min read🧠 Deep dive

Original authors: Konstantin Chernyshev, Ekaterina Artemova, Viacheslav Zhukov, Maksim Nerush, Mariia Fedorova, Iryna Repik, Olga Shapovalova, Aleksey Sukhorosov, Vladimir Dobrovolskii, Natalia Mikhailova, Sergei Tilga

Original paper licensed under CC BY 4.0 (http://creativecommons.org/licenses/by/4.0/). This is an AI-generated explanation of the paper below. It is not written or endorsed by the authors. For technical accuracy, refer to the original paper. Read full disclaimer

Imagine you need a complex project done, like organizing a massive amount of data, writing a detailed report, or researching a specific topic. You have three main options: hire a human freelancer, use a fully automated AI robot, or use Tendem.

According to the paper, Tendem is a "hybrid" team. Think of it as a high-speed race car driven by a human expert. The AI is the engine that does the heavy lifting, the fast driving, and the repetitive work. The Human Expert is the driver who keeps their eyes on the road, steers when things get tricky, and hits the brakes if the AI starts to hallucinate or get lost.

Here is how the paper breaks down this system in simple terms:

1. How It Works: The "Co-Pilot" System

The paper describes Tendem as a workflow where the AI and a Human work together, but they have specific roles:

  • The AI (The Engine): It handles the boring, fast, and repetitive stuff. It browses the web, reads files, and drafts content. It's like a super-fast intern who never sleeps.
  • The Human Expert (The Co-Pilot): They don't do every single step. Instead, they act as a safety net. They check the AI's work at specific "gates" (like checkpoints in a race). If the AI is unsure, or if the task is risky, the Human steps in to fix it, verify facts, and make sure the final product is perfect.
  • The Quality Check: Before the result is sent to the client, it goes through a double-check system. First, automated tools scan for errors. Then, a human expert gives it a final look to ensure it actually makes sense and meets the client's needs.

2. The Big Race: Tendem vs. The Competition

The researchers tested Tendem against two other groups using 94 real-world tasks (things like collecting business contacts, building automation workflows, or writing market research reports).

  • The Opponent A (Human Freelancers): These are people hired on platforms like Upwork. They are smart and reliable, but they can be slow and expensive.
  • The Opponent B (AI-Only): This is a fully automated AI (like a standard ChatGPT agent) trying to do the whole job alone. It's fast and cheap, but it often makes mistakes, misses details, or invents facts.

The Results:

  • Quality: Tendem won the gold medal. 74.5% of Tendem's results were rated "Good" (ready to use immediately).
    • Human freelancers got 53.2% "Good."
    • The AI-only bot only got 40.4% "Good."
    • Analogy: If you ordered 100 pizzas, Tendem would get 75 of them right. The human-only team would get 53 right, and the AI-only team would burn or undercook 60 of them.
  • Speed: Tendem was much faster than humans. It cut the total time in half (53% faster).
    • Humans took about 35 hours on average.
    • Tendem took about 16.5 hours.
    • Analogy: Tendem is like a delivery driver who uses a drone to drop off the package instantly, while the human freelancer has to drive a truck through traffic.
  • Cost: Tendem was cheaper than humans on average (a 36% reduction in median cost), though the AI-only version was the cheapest (but the quality was poor).

3. Why Did Tendem Win?

The paper suggests that Tendem wins because it fixes the weaknesses of both sides:

  • AI is fast but brittle: When AI gets confused or faces a complex, multi-step problem, it tends to make up facts or miss instructions. Tendem's human experts catch these errors before they reach the client.
  • Humans are careful but slow: Humans are great at judgment but get tired and take a long time to do repetitive data entry. Tendem's AI handles the boring stuff so the human can focus on the important decisions.

4. The "Brain" Behind the System

The paper also tested the AI part of Tendem on its own (without the human). They found that the AI is very strong at:

  • Web Browsing: Finding information online.
  • Tool Use: Using software to get things done.
  • Reasoning: Solving hard logic puzzles.

It's nearly as good as the best AI models available today. This proves that the "engine" is powerful enough to do most of the work, which is why the human only needs to step in occasionally to steer the ship.

Summary

Tendem is a "best of both worlds" solution. It combines the speed and low cost of AI with the reliability and judgment of a human expert. In the tests, it produced higher-quality work, faster, and for less money than hiring a human alone, while avoiding the mistakes that happen when AI works alone.

Note: The paper specifically states that this system is for general professional tasks (like research, data, and content) and does not cover high-stakes areas like medical advice, legal counsel, or senior-level programming.

Drowning in papers in your field?

Get daily digests of the most novel papers matching your research keywords — with technical summaries, in your language.

Try Digest →