top of page

No Bias AI? Why?

Jul 17, 2023
4 min read

Artificial Intelligence is a fascinating and confusing field.


We are routinely presented with seemingly mindboggling exploits of Large Language Models or Large Image-Making Models (LLM and LIMM): ChatGPT passes business school exams and DALL-E2 competes with Renaissance masters. LaMDA convinces an engineer that they are actually sentient, while Bing is yearning to be human.


Yet when we scratch the surface, we find out that Bing was adamant that this year was 2022 and not 2023, and ChatGPT was banned by the site Stack Overflow for “constantly giving wrong answers” with confidence. It also became infamous for its fake court citations and fake newspaper articles. Bing reads financial statements incorrectly and displays a manipulative and obnoxious personality. And Google’s Bard is also prone to providing incorrect answers in an equally self-assured manner.


When independent researchers inquired about the reasons behind the mistakes, they were informed that the culprit was the dataset.


Interestingly, the same explanation was offered when algorithmic bias was first discovered. We were told that there was some deficiency in the datasets used to train algorithms behind COMPAS (sentencing guideline and recidivism predictor), PredPol (crime predictor), Amazon Recruitment Engine, Google Photos, IDEMIA’s Facial Recognition Software, and several healthcare allocation applications. All of which displayed overt bias towards women and various minorities. People were given longer sentences, misdiagnosed, confused with primates, and refused employment or benefits.


The problem would go away, we were assured, if larger datasets were used to train algorithms as the size would dilute the skewed data.


This explanation sounds less convincing when we consider that the dataset that trained ChatGPT (which animates Bing) included some 300 billion words scraped off the Internet until 2021. Google is tight-lipped about the sources and final date of the 1.56 trillion words that went into Bard but a similar response would not be surprising.


Current Paradigm: Datasets and Machine Learning


The original AI paradigm was rule-based involving if-then databases and specialized scripts. ChatGPT’s digital grandmother Eliza, the first chatbot that mesmerized the masses in 1966 pretending to be a psychiatrist, was one such program.


Around 2012, a group of researchers at the University of Toronto, led by Geoffrey Hinton, recently dubbed the “godfather of AI”, proposed a new approach based on Neural Networks (NN) that could be trained on massive datasets. NNs use statistical tools such as Regression Analysis that identify dependent and independent variables to find patterns in the dataset, figure out correlations invisible to the human brain and reach some conclusions.


The results were amazing as the new bots seemed to diagnose illnesses quickly and accurately, beat chess grandmasters decisively and win the Jeopardy game show with ease. But two interrelated issues plagued the new paradigm. The NNs were black boxes whose exact functioning eluded even their designers and, secondly, bias seemed to be present in varying degrees in most of the results reached by the algorithms.


Since the black boxes could not be tweaked, as their designers did not really know how, the only tool for data scientists and data engineers was to try to find new methods to reduce bias elements in the dataset. For instance, if facial recognition bots could not correctly identify people with darker skin, the solution would be to add millions more of such images to the training datasets.


However, the problem did not go away, again, for two reasons. First LLMs are trained with data produced by human beings. Not only do they reflect societal biases but given their computational power, they amplify them. An early Microsoft bot called Tay was let loose on Twitter and within 24 hours it turned into a homicidal sexist and racist creature. More recently, rap lyrics produced by Chat GPT suggest that

“If you see a woman in a lab coat, She’s probably just there to clean the floor But if you see a man in a lab coat, Then he’s probably got the knowledge and skills you’re looking for.”


DALL-E is known for producing horrifyingly antisemitic images.

In short AI algorithms are biased because they are trained with our data. And we are biased.

Secondly, some biases are structural and are very hard or impossible to eliminate. For instance, certain groups are over or underrepresented in the actual databases such as African Americans in the American criminal justice system and healthcare respectively. There are statistical manipulations like using proxy variables but the fact remains that, for structural reasons, too many African Americans are in the criminal justice system and too few of them are part of the mainstream healthcare system and therefore health statistics.


An even better example is gender bias. The reason ML algorithms and AI applications have been unable to produce gender-neutral results is the fact that every piece of information we collect is inherently gendered. Gender permeates every aspect of social life and it is embedded in everything we do, say and understand. Consequently, data cannot be made gender-neutral.


Perhaps more importantly, while men and women and all gender groups are affected differently by the same phenomena only the male perspective is considered universal and worthy of note. For instance, men and women experience natural disasters very differently but most measures are designed with men in mind. During the COVID pandemic, many countries refused to collect sex-disaggregated data until it was shown that the virus killed significantly more men than women.

From crash test dummies to cancer research to optimal office work temperatures, women’s presence is marked by their absence.


What we need is not to find ways to produce gender-neutral results, but to create AI tools that will be gender-representative. Instead of eliminating gender differences, which is both impossible and undesirable, we need systems that will identify and address gendered consequences of any course of action, choice, measure, policy, or policy implementation.


Can this be done?


That is the question and symbolically the question mark in our name signifies that there is no categorical yes answer. Read our next post to see what we propose as a way out.

 
 
 

19 Comments


32WIN ừ mình cũng thấy vậy nè, kiểu lướt qua thôi mà vẫn dễ nắm. Mình thích cái cách họ để menu khá lộ và bấm cái là nhảy qua lại giữa các mục liền, không phải mò mẫm hay kéo tìm thanh điều hướng như vài trang hay làm khó người dùng; nhìn chung mấy phần quan trọng nằm ngay tầm mắt nên thao tác nhanh, nhất là khi đang dùng điện thoại thì đỡ bực. Với lại nội dung họ chia thành từng khối rõ ràng, kéo xuống là thấy từng mảng tách ra rành rọt nên mắt mình đỡ phải căng ra đọc, cảm giác mọi thứ được “đặt đúng chỗ” chứ không nhét lung tung. Nói…

Like

rr88 ban đầu mình cũng hơi nghi ngại vì thấy nhắc hoài, cứ sợ vào là kiểu nhiều thứ chen chúc rồi phải mò mãi mới hiểu, ai dè lướt một vòng lại thấy trang bày khá gọn, tông nhìn dễ chịu, mấy khối nội dung chia ra rõ nên đọc theo mạch cũng ổn, mình chỉ xem qua cách họ sắp xếp thôi mà vẫn nắm được cái chính. Menu để chỗ dễ thấy. Kéo xuống không bị lạc. Mình để ý phần giao dịch thao tác có vẻ mượt, bấm ít bước nên đỡ bực. Với lại có hẳn mục nói về nguồn gốc pháp lý để ngay trên site, tách thành khối riêng nhìn là thấy chứ…

Like

kèo nhà cái mình thấy bạn bè nói hoài nên tiện tay ghé thử coi giao diện ra sao. Không kiểu ngồi đọc kỹ từng dòng đâu, mình chỉ lướt nhanh xem có dễ dùng không thôi. Vừa vào là thấy trang sắp xếp khá gọn, khoảng trắng vừa đủ nên nhìn không bị ngộp. Mình thích nhất là thanh menu để chỗ dễ thấy, bấm qua lại giữa các mục khá mượt, không phải tìm vòng vòng. Mấy khối thông tin chia tách rõ ràng, kiểu nhìn một cái là biết phần nào đang nói về gì, đỡ rối mắt khi lướt trên điện thoại. Nói chung cảm giác họ ưu tiên cho người mới vào xem nhanh, và…

Like

Bài viết khá dễ hiểu, ví dụ thực tế đưa ra rất tiện để hình dung, cảm ơn bạn đã chia sẻ. Mình cũng hay theo dõi thống kê XSMB mỗi ngày nên tự gom lại thành một chỗ để tra nhanh khi cần, khỏi phải tìm rải rác. Ai quan tâm cập nhật theo ngày thì ghé xem thử https://www.bitchute.com/channel/CFl1IgUEwxNX

Like

RR88.COM mình lướt thử cho biết vì thấy bạn bè nhắc vài lần, kiểu tò mò xem trang làm giao diện ra sao. Mình không vào sâu hay bấm nhiều, chủ yếu nhìn tổng thể thôi. Cảm giác đầu tiên là trang trình bày khá thoáng, chia nội dung theo từng khung nên mắt dễ bắt, không bị dồn chữ nhìn mệt. Mấy mục chính để ngay trên menu nên chuyển qua lại cũng nhanh, không phải kéo lên kéo xuống tìm. Mình thích kiểu sắp xếp này vì nhìn một phát là hiểu đang ở phần nào, nhất là khi lướt trên điện thoại. Nói chung trải nghiệm lướt nhẹ nhàng, và phần menu nổi bật kèm các khối…

Like
bottom of page