PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 1, 20244 citationsOpen Access

Developing Safe and Responsible Large Language Models -- A Comprehensive Framework

View Full Paper
SRShaina RazaOBOluwanifemi BamgboseSGShardul Ghuge

Key Points

Key points are not available for this paper at this time.

Abstract

Given the growing concerns around the safety and risks of Large Language Models (LLMs), it is essential to develop methods for mitigating these issues. We introduce Safe and Responsible Large Language Model (SR₋₋₌), a model designed to enhance the safety of language generation using LLMs. Our approach incorporates a comprehensive LLM safety risk taxonomy and utilizes a dataset annotated by experts that align with this taxonomy. SR₋₋₌ is designed to identify potentially unsafe content and produce benign variations. It employs instruction-based and parameter-efficient fine-tuning methods, making the model not only effective in enhancing safety but also resource-efficient and straightforward to adjust. Through our testing on five benchmark datasets and two proprietary datasets, we observed notable reductions in the generation of unsafe content. Moreover, following the implementation of safety measures, there was a significant improvement in the production of safe content. We detail our fine-tuning processes and how we benchmark safety for SR₋₋₌ with the community engagement and promote the responsible advancement of LLMs. All the data and code are available anonymous at https: //github. com/shainarazavi/Safe-Responsible-LLM.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Raza et al. (2024) studied this question.

synapsesocial.com/papers/68e7101bb6db643587688f88https://doi.org/10.48550/arxiv.2404.01399
Ask AI
Helpful
Bookmark
Share
View Full Paper