Norvin scores live calls for synthetic speech and hands your team an evidence-backed verdict in about six seconds — which models agreed, at which moments, and why. It recommends verification. It never blocks or terminates a call.
Commercial text-to-speech now passes casual human inspection. The channels where money and identity move — contact centers, third-party verification, telephone payments, help desks — have no instrument for the question "is this voice generated?" Every workflow that treats a convincing voice as evidence is quietly exposed.
Account-ownership changes, identity-recovery resets, payment authorization, SIM-swap and activation flows, TPV consent capture. The calls where a convincing voice is treated as proof.
Thresholds tuned on clean studio audio are measurably wrong on 8 kHz telephony. Norvin is calibrated per condition — clean and telco paths each get their own measured operating point.
Advisory verdicts, human override, per-tenant kill switch, full audit trail. A detection system that can prove what it does — and cannot act alone.
calls with embedded synthetic speech flagged, controlled evaluation
† Controlled evaluations on partitioned datasets with frozen operating points; full populations, sample sizes, and Wilson confidence intervals in the methodology. Field false-positive rates are listening-adjudicated; raw flags containing real machine speech are reported separately.
Telephony conditioning breaks the two strongest detector architectures in opposite directions. The measurements that made a third family non-optional.
Read →RESEARCHNeither synthetic codec conditioning nor real PSTN transmission reproduced our false positives. Spontaneous speech did. Domain gaps live in speech style, not codec math.
Read →TECHNICALWhy checkpoints selected on benchmark metrics produced telephony false positives, and what selecting on deployment audio looks like.
Read →