VOICE TRUST · ADVISORY BY DESIGN

Synthetic-voice detection with the evidence attached.

Norvin scores live calls for synthetic speech and hands your team an evidence-backed verdict in about six seconds — which models agreed, at which moments, and why. It recommends verification. It never blocks or terminates a call.

NEVER IN THE CALL PATHFAIL OPENHUMAN OVERRIDE
CONSOLE · CALL DETAIL · CALL_ID HARNESS-8F3CILLUSTRATIVE
FUSED RISK TIMELINE · 4S WINDOWS
0swindow 08 · high band ▸80s
HIGH BAND · CONFIDENCE 0.91
advisory · recommends step-up verification
FLAGGED
PER-FAMILY SCORES
family A · clean-path
0.31
family B · high-recall
0.88
family C · adjudicator
0.79
REASON CODES
cadence-regularityspectral-artifactprosody-flatformant-drift
Norvin operator console — fused per-window verdicts with reason codes. Illustrative rendering, seeded with an evaluation-harness call.
01THE PROBLEM

Voice used to authenticate itself. That's over.

Commercial text-to-speech now passes casual human inspection. The channels where money and identity move — contact centers, third-party verification, telephone payments, help desks — have no instrument for the question "is this voice generated?" Every workflow that treats a convincing voice as evidence is quietly exposed.

01

High-consequence workflows

Account-ownership changes, identity-recovery resets, payment authorization, SIM-swap and activation flows, TPV consent capture. The calls where a convincing voice is treated as proof.

02

Lab detectors fail on the phone network

Thresholds tuned on clean studio audio are measurably wrong on 8 kHz telephony. Norvin is calibrated per condition — clean and telco paths each get their own measured operating point.

03

Detection you can govern

Advisory verdicts, human override, per-tenant kill switch, full audit trail. A detection system that can prove what it does — and cannot act alone.

02MEASURED, WITHIN STATED BOUNDS

Every number states its dataset, sample size, and confidence interval.

PARTIAL-SYNTHETIC DETECTION
19/20

calls with embedded synthetic speech flagged, controlled evaluation

TIME TO FIRST VERDICT
~6s

from call start, fused across three families

HELD-OUT TTS VOICES
100%

flagged at the monitoring band, telco condition, n=240

COMPUTE
CPU-only

commodity hosts, no GPU; measured unit economics

† Controlled evaluations on partitioned datasets with frozen operating points; full populations, sample sizes, and Wilson confidence intervals in the methodology. Field false-positive rates are listening-adjudicated; raw flags containing real machine speech are reported separately.

03RESEARCH

From the lab.

ALL RESEARCH →
04HOW WE START

Shadow-first. A slice of your traffic. Walk away anytime.