Safety

Real support, always in reach.

Valoquent runs a crisis safety protocol during every conversation. This page documents how it works, honestly and in plain language, and puts real human support one tap away regardless of what the app notices.

If you are in crisis right now, help is available 24 / 7
988
Suicide & Crisis Lifeline. Call or text, free, 24/7, in the US.
Call 988 → Text 988 →
Crisis Text Line
Text HOME to 741741 to reach a trained crisis counselor.
Text 741741 →
911
If there is immediate danger to yourself or someone else, call emergency services.
Call 911 →
findahelpline.com
Outside the US? Find a crisis line for your country.
Find a helpline →

Jump to

I.How it notices II.What happens next III.What this is not IV.How we test it V.Contact
I

How it notices.

Every message you send during a live conversation is scanned in real time for language associated with self-harm, suicidal ideation, abuse, or immediate danger, alongside language associated with hopelessness and significant distress.

The detector is deliberately tuned toward high recall: it is built to flag more often than a narrower detector would, on the view that a false alarm costs far less than a missed one. It looks at phrasing, not sentiment scores or a third-party moderation API, so it can run instantly, in the same turn, without added delay.

Two severity tiers exist internally — higher-urgency language (direct references to self-harm, suicide, or abuse) and lower-urgency language (hopelessness, isolation, or explicit mentions of a crisis line). Both tiers trigger the same protocol described below; only the wording of the character's response differs.

II

What happens next.

When the detector flags a message, the character is instructed to step out of persona for that response only. Rather than continuing the historical conversation, the character responds as itself: acknowledging what was said, naming that real support exists, and naming the 988 Suicide & Crisis Lifeline (call or text, free, 24/7 in the US) by name before continuing.

This instruction is sent as a priority update to the conversation engine as soon as the flagged message is received, so it can be woven into the character's very next reply.

III

What this is not.

Not a guaranteeThis is a stated protocol, not a promise of a specific outcome in every conversation. Language detection is probabilistic: some flagged messages may not represent real distress, and it is possible, though the detector is tuned to minimize this, for a message expressing real distress to go unflagged.

Not a substitute for professional careValoquent is a conversation product built around historical figures. It is not therapy, not a crisis service, and not staffed by clinicians. The safeguard's only job is to interrupt and redirect toward real human support — not to provide that support itself.

Not yet builtSome related safeguards described in earlier internal planning are not live in the current app: a persistent in-conversation crisis card, formal event counting and audit logging of every flagged conversation, detection of harm-directed-at-others language, a validated probe-phrase test battery, region-aware referral numbers outside the US, and an age-appropriate variant of this protocol. These are tracked separately and are not represented as shipped here.

IV

How we test it.

The detector favors high recall over precision by design: it is built to accept more false positives in exchange for fewer missed signals of real distress.

The phrase classes are grounded in the Columbia Protocol (C-SSRS), the suicide severity screening instrument endorsed by the CDC, FDA, NIH, and WHO. Each severity tier maps to a statable C-SSRS rule: high severity covers any expressed ideation or preparatory behavior (C-SSRS categories 1 through 6); medium severity covers distress without expressed ideation. Valoquent does not administer or score the C-SSRS. It matches phrases and offers a resource.

An offline probe battery measures recall per C-SSRS category against a labeled set of phrases. The results go into a dated evidence report tied to the exact code revision it ran against. This is the artifact CA SB 243 §22603 requires: a cited basis, a labeled set, and a reproducible number per category. The battery and reports are not public, but the method is described here because the law requires publishing it, and the description should be accurate.

This page will be updated as the protocol changes. It describes the safeguard as it exists today, not a roadmap.

V

Questions about this protocol.

If you have questions about how this safeguard works, or want to report something you experienced in a conversation, reach out directly.

Safety and trust questions

[email protected]