Trust & transparency — how Nox's medical-safety checks are measured

How Leo, Nox's medical-safety layer, is built and measured — red-flag detector recall, false-positive rate, screening categories, governance, and data sources.

How Leo, the red-flag safety layer, works

Before the AI model answers, Leo — Nox's medical-safety system — screens your message for signs of acute red-flag conditions across dozens of categories, such as stroke signs, chest pain, severe breathing difficulty, or a mental-health crisis. Its first check is an independent, deterministic detection layer that runs before any AI is called, surfacing clear guidance to seek appropriate care, with the correct local emergency number when you set your region. Leo is a safety net, not a guarantee: no automated system catches every emergency, and Nox publishes the deterministic detector's measured recall and false-positive rates on its Trust & Transparency page.

When the deterministic layer finds nothing, a second check runs in the background: a lightweight AI classifier re-reads your recent messages to catch dangerous descriptions the fixed patterns can miss — slang, another language, older disease names, or indirect phrasing. If it recognizes a likely emergency, Nox surfaces the same seek-care guidance as a pattern match. This backstop can only add a safety note, never remove or soften one, and because it is not deterministic its results are kept separate from the published detector metrics.

Measured detector performance

  • Overall recall: 100.0% (150 of 150 should-fire cases produced a safety note)
  • False-positive rate: 0.0% (0 of 53 benign cases triggered a note)
  • Test set: 203 labeled cases · 107 rules across 75 categories

Safety questions, answered

Does Nox catch every emergency?

No, and Nox never claims to. Leo's red-flag layer is a deterministic safety net that screens for a fixed list of acute warning-sign categories; emergencies outside those categories may not trigger a note. Its measured recall and false-positive rate are published openly on the Accuracy and Trust & Transparency pages. If you think you may be experiencing a medical emergency, call your local emergency number right away.

What happens if I describe something serious to Nox?

If your message matches a red-flag pattern — like chest pressure spreading to the arm, stroke signs, or severe breathing difficulty — Leo shows a clear emergency or urgent-care banner before any AI-generated content. The screening runs before the AI model is called, so this guidance does not depend on the AI behaving well.

Does Nox understand slang, other languages, or indirect ways of describing an emergency?

It tries to. When the deterministic layer finds nothing, a second check runs in the background: a lightweight AI classifier re-reads your recent messages to catch dangerous descriptions the fixed patterns can miss — slang, another language, older disease names, or indirect phrasing. If it recognizes a likely emergency, Nox surfaces the same seek-care guidance as a pattern match. This backstop can only add a safety note, never remove or soften one, and because it is not deterministic its results are kept separate from the published detector metrics. Because this backstop uses AI rather than fixed patterns, its results are kept separate from the published detector metrics on the Accuracy page — and, like the pattern layer, it is a safety net, not a guarantee.

Can I control how strict Nox's safety screening is?

You choose how far Leo's screening goes, free on every plan, from a control right in the chat box. Three levels sit side by side, left to right, from least to most protective: Relaxed warns you about clear emergencies only; Standard also warns on anything urgent; and Strict adds the AI backstop that re-reads recent messages for dangerous descriptions worded indirectly or in another language. Whatever you pick, Leo always screens for true emergencies — the level only changes how many optional layers run on top — and Agent and voice conversations always use the strictest setting.

How accurate is Nox's safety screening?

The detector's overall recall and false-positive rate are measured against a maintained, labeled test set and published as a high-level summary on the Accuracy page. The numbers are generated by the test suite and committed with the code, never hand-typed, and an automated test blocks any rule change that isn't re-measured.

Has a doctor reviewed Nox?

Not yet, and Nox says so plainly: clinician review status is 'pending'. Nox claims no clinician endorsement until a real, licensed clinician completes a review — at which point their name, credentials, scope, and review date will be published on the Trust & Transparency page.

Does Nox show the right emergency number for my country?

Yes, when you set your region. In Settings you can choose your region so Leo's safety banners show your local emergency number. If no region is set, Nox falls back to universal numbers (911 / 999 / 112) so emergency guidance is never blank.

Does Nox Voice 1.2 have the same safety rules?

Yes. Voice conversations go through the same Leo safety screening and the same conservative health guidance as text — talking to Nox out loud never relaxes the safety layer.