Key points are not available for this paper at this time.
Users report prejudiced responses generated by large language models (LLMs) like ChatGPT. Across 3 preregistered experiments, members of stigmatized social groups (Black Americans, women) reported higher trustworthiness of LLMs after viewing unbiased interactions with ChatGPT compared to when viewing AI-generated prejudice (i.e., racial or gender disparities in salary). Notably, higher trustworthiness accounted for increased behavioral intentions to use LLMs, but only among stigmatized social groups. Conversely, White Americans were more likely to use LLMs when AI-generated prejudice confirmed implicit racial biases, while men intended to use LLMs when responses matched implicit gender biases. Results suggest reducing AI-generated prejudice may promote trustworthiness of LLMs among members of stigmatized social groups, increasing their intentions to use AI tools. Importantly, addressing AI-generated prejudice could minimize social disparities in adoption of LLMs which might further exacerbate professional and educational disparities. Given expected integration of AI in professional and educational settings, these findings may guide equitable implementation strategies among employees and students, in addition to extending theoretical models of technology acceptance by suggesting additional mechanisms of behavioral intentions to use emerging technologies (e.g., trustworthiness). • Three experiments examine the effects of prejudice generated by LLMs. • AI-generated prejudice reduced trust among Black Americans and UK women. • Distrust subsequently diminished intentions to use LLMs. • Stronger implicit biases promoted intentions among White Americans and UK men.
Petzel et al. (Wed,) studied this question.