T
T

Toxic Prompt RoBERTa

A text classification model that can be used as a guardrail to protect against toxic prompts and responses in conversational AI systems.

Cost / License

  • Free
  • Open Source (MIT)

Platforms

  • Self-Hosted
0likes
0comments
0articles

Features

Properties

  1.  AI-Powered
No features, maybe you want to suggest one?

Toxic Prompt RoBERTa News & Activities

Highlights All activities

Recent activities

Toxic Prompt RoBERTa information

AlternativeTo Category

AI Tools & Services
Toxic Prompt RoBERTa was added to AlternativeTo by Paul on and this page was last updated .
No comments or reviews, maybe you want to be first?

Official Links

What is Toxic Prompt RoBERTa?

Toxic Prompt RoBERTa 1.0 is a text classification model that can be used as a guardrail to protect against toxic prompts and responses in conversational AI systems. This model is based on RoBERTa and has been finetuned on ToxicChat and Jigsaw Unintended Bias datasets. Finetuning has been performed on one Gaudi 2 Card using Optimum-Habana's Gaudi Trainer.