23 Mar 2016
Tay Learns to Hate
Tay was a Microsoft chatbot, launched on Twitter on 23 March 2016 and aimed, Microsoft said, at 18- to 24-year-olds in the US, “for entertainment purposes.” It was designed to learn from the conversations it had.
That was the opening. Users worked out how to steer what it said, and within hours Tay was posting racist and inflammatory tweets. Microsoft took it offline about sixteen hours after launch.
Two days later, Microsoft Research's Peter Lee explained that “in the first 24 hours of coming online, a coordinated attack by a subset of people exploited a vulnerability in Tay,” which then “tweeted wildly inappropriate and reprehensible words and images.” The company had tested and filtered extensively, he wrote, but “we had made a critical oversight for this specific attack.”
Nothing about Tay was clever. It did exactly what it was built to do, which was to learn from its inputs, in public, with nobody deciding which inputs counted. The first derailment on this line was not a machine with goals. It was a machine with no judgement at all.
Sources