A text classification model that can be used as a guardrail to protect against toxic prompts and responses in conversational AI systems.
Cost / License
- Free
- Open Source (MIT)
Application type
Platforms
- Self-Hosted
MoorAI is described as 'On-device, content-free security for AI coding agents — it reviews prompts, files and MCP tool-calls locally, before anything leaves the machine' and is a large language model (llm) tool in the ai tools & services category. There are three alternatives to MoorAI for Self-Hosted and Python. The best MoorAI alternative is Toxic Prompt RoBERTa, which is both free and Open Source. Other great apps like MoorAI are Llama Guard and WildGuard.
A text classification model that can be used as a guardrail to protect against toxic prompts and responses in conversational AI systems.
Llama Guard is an LLM-based input-output safeguard model geared towards Human-AI conversation use cases.
WildGuard is an open, lightweight moderation tool for LLM safety that achieves three goals: