A Reinforcement Learning-based Offensive semantics Censorship System for Chatbots

07/13/2022
by   Shaokang Cai, et al.
0

The rapid development of artificial intelligence (AI) technology has enabled large-scale AI applications to land in the market and practice. However, while AI technology has brought many conveniences to people in the productization process, it has also exposed many security issues. Especially, attacks against online learning vulnerabilities of chatbots occur frequently. Therefore, this paper proposes a semantics censorship chatbot system based on reinforcement learning, which is mainly composed of two parts: the Offensive semantics censorship model and the semantics purification model. Offensive semantics review can combine the context of user input sentences to detect the rapid evolution of Offensive semantics and respond to Offensive semantics responses. The semantics purification model For the case of chatting robot models, it has been contaminated by large numbers of offensive semantics, by strengthening the offensive reply learned by the learning algorithm, rather than rolling back to the early versions. In addition, by integrating a once-through learning approach, the speed of semantics purification is accelerated while reducing the impact on the quality of replies. The experimental results show that our proposed approach reduces the probability of the chat model generating offensive replies and that the integration of the few-shot learning algorithm improves the training speed rapidly while effectively slowing down the decline in BLEU values.

READ FULL TEXT
research
10/25/2022

A Streamlit-based Artificial Intelligence Trust Platform for Next-Generation Wireless Networks

With the rapid development and integration of artificial intelligence (A...
research
02/09/2021

Security and Privacy for Artificial Intelligence: Opportunities and Challenges

The increased adoption of Artificial Intelligence (AI) presents an oppor...
research
04/20/2021

Prospective Artificial Intelligence Approaches for Active Cyber Defence

Cybercriminals are rapidly developing new malicious tools that leverage ...
research
01/30/2006

Instantaneously Trained Neural Networks

This paper presents a review of instantaneously trained neural networks ...
research
05/10/2019

Integrating Artificial Intelligence into Weapon Systems

The integration of Artificial Intelligence (AI) into weapon systems is o...

Please sign up or login with your details

Forgot password? Click here to reset