prompt_guard | Jailbreaks / prompt injections, scored by the NeuralTrust Firewall. | input, output | all | — |
toxicity | Toxic content, scored by the NeuralTrust Firewall. | input, output | all | — |
prompt_moderation (Moderation) | Off-topic / disallowed content via keyword+regex and/or NeuralTrust topics. | input, output | all | — |
url_analyzer | Fetches URLs in the content (SSRF-guarded) and screens fetched text for indirect prompt injection and PII. | input | llm, mcp | — |
doc_analyzer | Extracts text from uploaded documents (incl. OCR) and screens for PII and indirect prompt injection. | input | llm | — |