Audio deepfake detection.
Detect synthetic, cloned, or replayed speech and preserve evidence for review.
Audio review
Incoming call · CALL-0082
How it works.
Prepare the speech
Apply the documented channel and preprocessing policy.
Inspect the signal
Evaluate relevant segments at the operating threshold.
Preserve evidence
Connect the result with its segment and model record.
Route the outcome
Continue, step up, or send the call for review.
A cloned voice can clear a voiceprint. It still has to pass the authenticity check.
Check both speech authenticity and the enrolled speaker before advancing a sensitive request.
#1 in commercial latency.
24 ms per clip, with 97.3% of deepfakes detected. Results from the Podonos benchmark across 23 systems.
- Latency per clip
- 24ms#1 among commercial systems
- Processing speed
- 2.3×the next-fastest commercial system
- Deepfakes detected
- 97.3%of synthetic audio in the test
- Files scored
- 4,524Every file. None declined.
- Detection accuracy
- 94.47%#9 of 23 systems
Speed and detection accuracy
Higher is more accurate. Further right is faster.
Processing speed · 1× = real time
Commercial latency ranking
Average time per audio clip. Lower is faster.
- 1Detectif.ai24 ms
- 2Pella Research57 ms
- 3NII Synthetiq Audio91 ms
- 4Corsound AI180 ms
- 5Pindrop282 ms
- 6Resemble DETECT-World399 ms
- 7Fennura553 ms
- 8Hive881 ms
- 9Whispeak1.1 s
- 10Resemble AI1.2 s
- 11Reality Defender5.7 s
Reported timings · bars use a log scale
Accuracy scored by Podonos. Detectif.ai timings are vendor-reported; hardware and service conditions vary. The latency ranking includes the 11 commercial systems with published timing.
Podonos source · September 2026Built for the operating decision.
Relevant evaluation
Test the synthesis conditions that matter to the call flow.
Segment review
Preserve the speech associated with a review state when permitted.
Benchmark transparency
Keep metrics, provenance, and limitations visible.
Plan the deployment.
Detection performance depends on the dataset, channel, language, generator, threshold, and operating conditions disclosed with each evaluation.
Talk to the team- Which channels and codecs are expected?
- Which languages and generators matter?
- What false-positive cost can the flow tolerate?
- Which evidence may be retained?
Test it on your own call traffic.
A 30-minute walkthrough of your highest-risk call flow, then an evaluation on audio from your own lines.