this post was submitted on 07 Aug 2023
61 points (100.0% liked)

Technology

37362 readers
232 users here now

Rumors, happenings, and innovations in the technology sphere. If it's technological news or discussion of technology, it probably belongs here.

Subcommunities on Beehaw:


This community's icon was made by Aaron Schneider, under the CC-BY-NC-SA 4.0 license.

founded 2 years ago
MODERATORS
 

From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP Models https://aclanthology.org/2023.acl-long.656.pdf

you are viewing a single comment's thread
view the rest of the comments
[โ€“] [email protected] 8 points 11 months ago (1 children)

Despite what you might assume, an encyclopedia wouldn't be free from bias. It might not be as biased as, say, getting your training data from a dump of 4chan, but it'd absolutely still have bias. As an on-the-nose example, think about the definition of homosexuality; training on an older encyclopedia would mean the AI now thinks homosexuality is a crime.

[โ€“] [email protected] 4 points 11 months ago

And imagine how badly most encyclopedias would reflect on languages and cultures other than the one that made them.