Skip to main navigation Skip to search Skip to main content

Assessing Bias in Large Language Models: A Comparative Study of ChatGPT, Gemini, and Claude

Research output: Chapter in Book/Report/Conference proceedingConference contribution

Abstract

As Large Language Models (LLMs) become more integrated into various applications, their potential to display and amplify biases is a significant concern. This study evaluates the social and gender bias in ChatGPT, Gemini, and Claude AI. Unlike previous studies using pre-existing datasets, we used handcrafted questions to reveal biases more effectively. Basic simple questions provided an unbiased answer, however complex biased questions highlighted a few concerns. ChatGPT had a bias all the time while Gemini had a bias sometimes and the rest of the time it did not, or if it did the bias was different, while Claude AI was mostly unbiased. Thus, this study shows that there is variation in levels of bias in different AI models and thus advocates for enhanced strategies of Bias in AI to address the issue of fairness in the use of AI.

Original languageEnglish
Title of host publicationIEEE International Conference on Modeling, Simulation and Intelligent Computing, MoSICom 2024 - Proceedings
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages133-137
Number of pages5
ISBN (Electronic)9798331533311
DOIs
Publication statusPublished - 2024
Event2024 IEEE International Conference on Modeling, Simulation and Intelligent Computing, MoSICom 2024 - Dubai, United Arab Emirates
Duration: 09-12-202411-12-2024

Publication series

NameIEEE International Conference on Modeling, Simulation and Intelligent Computing, MoSICom 2024 - Proceedings

Conference

Conference2024 IEEE International Conference on Modeling, Simulation and Intelligent Computing, MoSICom 2024
Country/TerritoryUnited Arab Emirates
CityDubai
Period09-12-2411-12-24

All Science Journal Classification (ASJC) codes

  • Artificial Intelligence
  • Computer Science Applications
  • Energy Engineering and Power Technology
  • Electrical and Electronic Engineering
  • Modelling and Simulation
  • Instrumentation

Fingerprint

Dive into the research topics of 'Assessing Bias in Large Language Models: A Comparative Study of ChatGPT, Gemini, and Claude'. Together they form a unique fingerprint.

Cite this