Individual Submission Summary
Share...

Direct link:

Censorship’s Implications for Artificial Intelligence

Sun, September 13, 2:00 to 3:30pm MDT (2:00 to 3:30pm MDT), TBA

Abstract

While artificial intelligence provides the backbone for many tools people use around the world, recent work has brought attention to the potential biases that may be baked into these algorithms. While most work in this area has focused on the ways in which these tools can exacerbate existing inequalities and discrimination, we bring to light another way in which algorithmic decision making may be affected by institutional and societal forces. We study how censorship has affected the development of Wikipedia corpuses, which are in turn regularly used as training data that provide inputs to NLP algorithms. We show that word embeddings trained on the regularly censored Chinese language Wikipedia have very different associations between adjectives and a range of concepts about democracy, freedom, collective action, equality, and people and historical events in China than its uncensored counterpart Baidu Baike. We examine the origins of these discrepancies using surveys from mainland China and their implications by examining their use in downstream AI applications.

Authors