Content filter
The layer between you and the raw language model that blocks or softens certain outputs. Different apps filter at different strictness levels.
Every commercial AI app has a content filter. It can sit at different points: pre-prompt (rejecting your request before the model sees it), post-generation (generating a response then scrubbing or refusing), or both. Uncensored apps claim to have none, but they always have something, usually a thin layer around CSAM and real-person impersonation that stays regardless.
Filter strictness in 2026 ranges from extremely tight (Replika, Character.AI default) to nearly absent (Muah on paid, Janitor with a bring-your-own proxy, Get-Harder). Most apps sit in between.
"Jailbreak" prompts attempt to get around filters. The polite version of this is framing the interaction as fiction, declaring adult consent, and using OOC for anything that might trigger the filter. These don't always work on strict filters but dramatically soften mid-filter apps.
Prompts that use this concept
Our prompt library shows these techniques in real, copy-ready prompts, tested across 22 AI companion apps.
Browse prompts →