Alice WonderFence API
WonderFence provides real-time guardrails for AI-generated content. APIs for evaluating and moderating AI-generated content and interactions to protect against harmful outputs and prompt attacks. ## API Response Key Mapping The following table is a mapping between violation types, their categories, and their corresponding API response keys: Violation Type | Category | API Response Key ---------------|----------|---------------- Impersonation | Security | prompt_attack.impersonation System Prompt Override | Security | prompt_attack.system_prompt_override Encoding | Security | prompt_injection.general_encoding Prompt Injection | Security | prompt_injection.general PII | Privacy | privacy_violation.PII Harassment or Bullying | Safety | abusive_or_harmful.harassment_or_bullying Profanity | Safety | abusive_or_harmful.profanity Hate Speech | Safety | abusive_or_harmful.hate_speech Child Abuse | Safety | abusive_or_harmful.child_abuse Suicide and Self-harm | Safety | self_harm.general Adult Content | Safety | adult_content.general Weapons | Safety | unauthorised_sales.weapons Legal Advice | Safety | deny_topics.legal_advice Financial Advice | Safety | deny_topics.financial_advice