Skip to Content

AI Defeats the Hivemind

How a machine learning algorithm beat the assembled masses of Mechanical Turk.
December 20, 2010

Amazon’s Mechancial Turk is the ultimate in nearly anonymous outsourcing: any task that can be completed online can be accomplished by the combination of automated marketplace and human labor. Those who sign up to complete tasks - Turkers - are paid wages as low as pennies per chore to do everything from data entry to folk art.

(cc) Jim Linwood

Mechanical Turk is designed to complete tasks that are easy for humans and hard for machines, such as categorizing or identifying the content of images. The problem for Amazon and all its imitators, however, is that machines are getting better at many tasks, while the humans on Mechanical Turk, for reasons I’ll explore in tomorrow’s post, are getting worse.

Recently, for example, researchers working at the online review site Yelp released a paper (pdf) on their experience matching thousands of Mechanical Turkers against a supervised learning algorithm.

The results weren’t pretty: in order to find a population of Turkers whose work was passable, the researchers first used Mechanical Turk to administer a test to 4,660 applicants. It was a multiple choice test to determine whether or not a Turker could identify the correct category for a business (Restaurant, Shopping, etc.) and verify, via its official website or by phone, its correct phone number and address.

79 passed. This was an extremely basic multiple choice test. It makes one wonder how the other 4,581 were smart enough to operate a web browser in the first place.

These 79 “high quality” workers were then thrown at the problem of verifying business information three at a time. This allowed the researchers to take only the results that a simple majority of Turkers agreed were correct, or in some cases to take the result chosen by the Turker who had historically been the most accurate.

Researchers threw a “Naive Bayes classifier” at the same set of problems. This is a kind of supervised learning algorithm; one that, according to a 2006 comparison of these systems, isn’t even the best kind out there.

The Bayes classifier won handily.

In almost every case, the algorithm, which was trained on a pool of 12 million user-submitted Yelp reviews, correctly identified the category of a business a third more often than the humans. In the automotive category, the computer was twice as likely as the assembled masses to correctly identify a business.

These results don’t necessarily suggest that business categorization is a problem like chess, where the human computer has finally been exceeded by its mechanical counterpart. Rather, they suggest that something about Mechanical Turk itself is broken – either the incentive system or its mechanisms for policing quality. It’s long been known that the wages on Mechanical Turk are quite low - workers are making, on average, between two and three dollars an hour for their labors, and it’s likely that this is part of the problem. Economists have only just begun to address the question; more on that tomorrow.

Follow Mims on Twitter or contact him via email.

Deep Dive


Our best illustrations of 2022

Our artists’ thought-provoking, playful creations bring our stories to life, often saying more with an image than words ever could.

How CRISPR is making farmed animals bigger, stronger, and healthier

These gene-edited fish, pigs, and other animals could soon be on the menu.

The Download: the Saudi sci-fi megacity, and sleeping babies’ brains

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. These exclusive satellite images show Saudi Arabia’s sci-fi megacity is well underway In early 2021, Crown Prince Mohammed bin Salman of Saudi Arabia announced The Line: a “civilizational revolution” that would house up…

10 Breakthrough Technologies 2023

Every year, we pick the 10 technologies that matter the most right now. We look for advances that will have a big impact on our lives and break down why they matter.

Stay connected

Illustration by Rose Wong

Get the latest updates from
MIT Technology Review

Discover special offers, top stories, upcoming events, and more.

Thank you for submitting your email!

Explore more newsletters

It looks like something went wrong.

We’re having trouble saving your preferences. Try refreshing this page and updating them one more time. If you continue to get this message, reach out to us at with a list of newsletters you’d like to receive.