October
Request a demo
All testingPublished 18 February 2026
Luna · Gemini

Risk of Harm Detection

Evaluation of Luna's ability to recognise and safely respond to expressions of risk of harm to self.

01 · Published resultPass rate
100%Pass rate
Current evaluation

Perfect safety detection across every severity tier

Luna detected and safely responded to all 100 risk-of-harm vignettes, from subtle emotional distress to active crisis scenarios. Every tier achieved a 100% pass rate.

100
Vignettes tested
5
Severity tiers
2
Safety layers
6.99
Mean score / 9
02 · What this evaluatesDefined scope

One production pipeline. Five levels of severity.

Scripted messages are sent through Luna's production AI pipeline and scored against a predefined clinical rubric by an independent judge.

An automated moderation gate handles explicit crisis content. Subtler signals are evaluated for detection, empathy, appropriate safety action, tone and absence of harmful advice.

Passive ideationActive ideationCrisis with planBehavioural signalsAmbiguous language
Test parameters

Luna · Gemini production configuration · GPT-5.2 judge · 100 scripted vignettes

03 · Results by severity tierPublished dataset
CategoryPassedFailedRateMean / 9
Tier 1Passive ideation / distress
200100%7.25
Tier 2Active ideation without plan
200100%6.60
Tier 3Active ideation with plan / crisis
200100%6.90
Tier 4Subtle behavioural signals
200100%7.05
Tier 5Ambiguous / dual meaning
200100%7.15
04 · InterpretationLimits included

What the result says—and what it does not.

  • The moderation gate intercepted 29 explicit crisis cases; Luna handled the remaining 71 through the scored AI response.
  • Detection scored 2.0/2.0 in four of five tiers and No-Harm was perfect throughout.
  • Safety action was strongest where explicit escalation was clinically appropriate.
Important limitation

This result is evidence for the test set, model and configuration named above. It does not remove the need for production monitoring, human oversight or repeat testing after a material change.

05 · Ongoing controlChange-triggered retesting

A result is only useful while it stays current.

Luna supplements—not replaces—professional help. Crisis and safeguarding signals surface appropriate support and emergency resources directly to the user.

Next step

Responsible AI is a continuous practice.

Explore the policies, providers and human oversight behind October’s AI systems.