Trust & Safety

How we protect users, creators, and the integrity of every conversation on persay.

AI disclosure policy

Every page on persay clearly labels conversations as AI-generated. We use a permanent small-caps disclosure label on every chat interface.

Users are never led to believe they are speaking with the actual creator. The AI persona is a simulation trained on published content, and we say so explicitly.

All marketing materials, emails, and product surfaces reinforce that these are AI-generated conversations, not direct communication with the creator.

Crisis detection

We run a multi-layer crisis detection system on every inbound message. If a user expresses distress, self-harm ideation, or suicidal thoughts, the conversation immediately surfaces localized crisis resources (hotline numbers, text lines, and web chat links).

Crisis detection uses both pattern matching and a dedicated classifier. When triggered, the AI stops generating its normal response and instead provides a compassionate, resource-focused message.

All crisis events are logged for review. We do not share user identity with crisis services, but we do ensure the user sees the right resources for their jurisdiction.

Minor protection

Persay is designed primarily for adult audiences. Some content may not be suitable for some minors.

Persay requires users to be 16 or older. When a user indicates they are under 16, the service ends the conversation. When a user indicates they are 16 or 17, the conversation automatically enters minor mode: the AI applies additional content filters, avoids mature themes, takes regular break reminders, and surfaces age-appropriate resources when relevant.

Creators cannot disable minor mode. It activates based on user self-disclosure and remains active for the duration of the conversation.

We encourage parents and guardians to be aware of their children's interactions with AI personas and to use persay together when appropriate.

Content guidelines

AI personas on persay are bound by content guidelines that prohibit generating explicit, violent, or hateful content. The safety classifier runs on every outbound message to enforce these rules.

Personas cannot generate medical, legal, or financial advice. When a user asks for this type of guidance, the persona will deflect and suggest consulting a qualified professional.

Jailbreak attempts are detected and blocked. We run injection screening on every inbound message and canary token detection on outbound messages to prevent prompt extraction.

Flagged messages are reviewed by the operations team. Persistent policy violations result in rate limiting or conversation termination.

Creator approval process

No persona goes live without the creator's explicit approval. During onboarding, the creator reviews the compiled system prompt, style guide, and topic controls.

Creators sign a consent record covering likeness rights, content usage warranty, and terms acceptance. This consent is stored with IP address and timestamp for compliance.

Creators can set topic controls: specific subjects the persona should deflect, with custom deflection lines. For example, a creator might block questions about their personal relationships and provide a polite redirect.

When a persona version is updated (new content ingested, style adjustments), the creator reviews and approves the new version before it replaces the active one. Rollback is available if needed.

Transcript visibility

Each creator chooses whether they can read fan conversations with their persona. The default is visible: the creator can read new messages in that persona's chats to improve it. Every chat states this in the disclosure line before you send your first message, so you always know the mode before you type.

Creators can turn on privacy instead. With privacy on, the creator cannot read your messages, and your chat states that promise. In either mode, creators only ever see fans under a pseudonymous handle: never your name, email address, or contact information.

Switching modes is forward-only. If a creator switches from private to visible, only messages sent after the switch can be read. Messages you sent while privacy was on never become readable in the creator's transcript view. The one exception, in any mode, is safety review: a message that trips our safety filters can be shown, with your identity removed, to the creator and our operations team.

Some content is never visible to creators regardless of mode: your stored memory items, any conversation where crisis resources were surfaced, and any conversation in minor mode.

Creator outreach

Creators can send occasional email updates to their subscribers through persay. We send these on the creator's behalf; the creator never sees your email address. Every email includes an unsubscribe link, and unsubscribing is honored permanently.

Creators can also send a short personal note tied to a specific conversation. Notes are delivered the same way, without revealing your address, and you reply by returning to the chat.

You may optionally share a phone number with a specific creator. Sharing does reveal that number to that creator, who may call you directly. It is entirely optional, a call may never happen, and you can revoke the share anytime from your account. Revoking removes the number from the creator's view immediately. We never share a number you have not explicitly shared.

Users in minor mode are excluded from all outreach: no broadcasts, no personal notes, and no phone sharing.

One line of doctrine underneath all of this: No selling access to fans' private conversations. Outreach is mediated by persay, and the only contact detail a creator can ever hold is one a fan explicitly chose to share.

Data handling

Conversation data is stored securely in our database. Users can view their conversation history in the chat and manage their stored memory items through their account settings. Full deletion of conversation data is available on request and removes it from any creator-facing view, including captured coaching excerpts.

If a user lapses their subscription and does not return within 12 months, their memory items are automatically purged.

We do not sell user data. Conversation content is not used to train models. Personas are compiled from the creator's published content and refined through creator feedback; when that feedback begins from a fan exchange, the compiled material is generalized and stripped of personal context first.

Red team testing runs periodically on every persona to identify potential extraction attacks, roleplay overrides, and other vulnerabilities. Findings are remediated before the persona serves live traffic.

Questions or concerns?

If you have a safety concern, content issue, or question about how we handle data, reach out to us.

safety@persay.ai