Truthring
Comparison · every claim about ElevenLabs checked 31 August 2026

Truthring vs the ElevenLabs AI Speech Classifier

The short answer

These two products answer different questions. Checked 31 August 2026, ElevenLabs’ own documentation calls the AI Speech Classifier “A legacy tool” and points to a newer watermark-based Audio Detector, and its FAQ states it is “not a general-purpose detector for every AI voice provider”. A generator checking its own output is establishing provenance — stronger evidence than an outside detector can produce, over a much narrower question. Truthring is built for the wider one: a clip of unknown origin.

We sell a competing product, so nothing here is neutral. What we offer instead is traceability: every statement about ElevenLabs below was read on one of their live pages on 31 August 2026 and carries the URL it came from.


What the classifier gets right, said first

It states the scope of a negative result, in its own words. Asked “Does it detect audio from other AI voice tools?”, the answer published at elevenlabs.io/ai-speech-classifier on 31 August 2026 reads: “No. The classifier is built to detect speech generated with ElevenLabs. It is not a general-purpose detector for every AI voice provider.” Most detectors leave the meaning of a clean result to the reader’s imagination. This one closes it in two sentences.

It names a case where it stops working, and the case is its maker’s own current model. The same page states that the classifier “Does not reliably classify audio generated with the Eleven v3 model”. A vendor publishing the model its own detector struggles with is the rarest disclosure in this category.

It dates the edge of its provenance signal. “No ElevenLabs audio created prior to June 2026 carries a watermark.” A date is falsifiable, and it tells you exactly which body of audio the newer detector cannot help with.

The page also publishes “99% precision, 80% recall” for unmodified ElevenLabs audio. Recall of 80%, stated openly, is more candour than a headline “99% accurate” ever contains. Beside those numbers we found no false positive rate, no dataset, no sample size and no test date.


What each of us publishes

Published itemTruthringElevenLabs AI Speech Classifier — checked 31 Aug 2026
Scope of the answer
What a clean result rules out
Synthetic or not, whatever the generatorElevenLabs speech only, stated in their FAQ: “No. The classifier is built to detect speech generated with ElevenLabs. It is not a general-purpose detector for every AI voice provider.” — ai-speech-classifier
Product statusPre-launch. No live detector, no customers — aboutElevenLabs’ own documentation calls it “A legacy tool” and points to a newer watermark-based Audio Detector. The older documentation URL returned 404 on 31 Aug 2026.
Headline numbersNone. No measured figure exists yet — accuracy“99% precision, 80% recall” for unmodified ElevenLabs audio.
False positive rate
Genuine speech called synthetic
Committed, not yet measured. Framework published in advance — accuracyNOT FOUND at ai-speech-classifier on 31 Aug 2026.
Dataset, sample size and test date behind the numbersTest-set design published before any run — evaluation datasetNOT FOUND at that URL on 31 Aug 2026.
Stated failure conditionsPublishedOne published, and specific: “Does not reliably classify audio generated with the Eleven v3 model.”
Provenance boundaryNot applicable. We read the audio; we place no mark at creationPublished: “No ElevenLabs audio created prior to June 2026 carries a watermark.”
Written, versioned methodPublished — version 2.4, datedNo methodology page found on 31 Aug 2026.
Model version on each resultYes — engine version on every verdictNOT FOUND on 31 Aug 2026.
Reference code for re-examinationYes — a Ringmark on every verdictNOT FOUND on 31 Aug 2026.
Named peopleYes — author profileNOT FOUND on 31 Aug 2026.
CertificationNone held. Pre-launch — compliance“We’re certified SOC2 and GDPR compliant” at elevenlabs.io/enterprise. Their /trust and /security URLs returned 404 on 31 Aug 2026.
Audio retention, in time unitsPeriods not set — pre-launch. The design is published, the numbers are not — data retentionNo classifier-specific period found. General policy states voice data is not kept “longer than 3 years after your last interaction”; Zero Retention Mode is enterprise-only.
Commercial interest in generationNone — we sell no synthesis of any kindSells cloning and synthesis — elevenlabs.io/voice-cloning
Legal entity publishedLacewing Technologies, Navi Mumbai, IndiaEleven Labs Inc., 169 Madison Ave #2484, New York NY 10016.

Right-hand column read on the linked pages on 31 August 2026. Re-checked quarterly. Tell us what is wrong and the correction is published with a date: corrections@aivoicedetctor.com


A generator checking its own output is doing provenance, not detection

Did ElevenLabs make this? On that question ElevenLabs is not guessing. They hold the model, can place a signal at generation and read it back, and may hold a record of what was produced. If a mark survives to the file you are holding, that is stronger evidence than any statistical reading of a waveform, ours included.

Did a machine make this? There a first-party tool is scoped out by its own FAQ, honestly. A clean result means not ours. It does not mean not synthetic. On screen the two look identical — a low number, a reassuring colour — and if the clip came from another vendor’s model, or an open-weights system someone ran on a laptop, “not ours” is entirely correct and entirely useless.

So this is a comparison of purposes, not of quality. Where the origin is known, the platform wins. Where it is unknown — nearly every recording anyone brings to a detector — its tool was never in the running, and says so itself.

The watermark boundary sharpens that. “No ElevenLabs audio created prior to June 2026 carries a watermark” means every ElevenLabs clip made before that date falls back on the tool the same documentation calls legacy. If the disputed voicemail on your desk is from last year, the stronger method has nothing to read. And one question worth putting to any watermark vendor, asked rather than asserted because we have not tested theirs: what survives re-encoding, a screen recording, or a trim to eight seconds?

The structure, described without accusation

A company that sells voice cloning and also grades its own output sits inside a real tension: a detector that catches its own product is, read one way, a report card on that product, and a vendor in that position has no commercial reason to become good at detecting a rival’s output.

Both halves point away from suspicion here. Nobody knows a system better than the people who built it, and ElevenLabs has already published what its scope is — so the disclosure that normally settles a structural question is already on their own website. We sell no synthesis, which is worth one thing only: no second product our verdicts could embarrass.


Where ElevenLabs is the better choice

  • Your question is specifically about ElevenLabs. A licensing dispute, a terms complaint, an account investigation. They answer from provenance; we would be inferring.
  • The audio was created on their platform after June 2026. The watermark detector then reads a mark placed at generation, which beats waveform analysis, ours included.
  • They publish two things we do not yet. A retention period in time units — voice data not kept “longer than 3 years after your last interaction” — and a certification claim. We have neither, because we have not launched.
  • Their speech engineering is not in question, and it would be silly to pretend otherwise.

Where we are the better choice

  • The clip could have come from anywhere. A scam call, a leaked recording, a voice note forwarded four times. By its own published scope, the classifier is not built for that.
  • The audio predates June 2026, or is not ElevenLabs at all. No watermark exists to read, so the question returns to inference — the only thing we do.
  • The answer has to survive being questioned. A verdict from a party with no stake in the generator, carrying a reference code and an engine version, is a different object in a dispute.
  • You want the failure modes written down, including when we return unclear: limitations.

Questions

Which is more accurate?

We are not going to rank it. Accuracy is comparable only when two detectors run on the same held-out clips at the same thresholds, judged by someone with no stake, and nothing like that exists for this pair. Truthring is also pre-launch, with no measured figure at all yet.

Does the classifier detect audio from other AI voice tools?

According to its own FAQ, read on 31 August 2026: no. The published answer is “No. The classifier is built to detect speech generated with ElevenLabs. It is not a general-purpose detector for every AI voice provider.”

Is the classifier still the tool ElevenLabs recommends?

Their documentation described it as a legacy tool on 31 August 2026 and pointed to a newer watermark-based Audio Detector. The watermark covers audio created from June 2026 onwards, so older ElevenLabs audio still falls to the legacy classifier.

Should I trust a comparison written by a competitor?

No — check it. Every cell in the right-hand column names the page it came from and the date it was read.

All claims about ElevenLabs verified 31 August 2026 against their own live pages, linked above. Re-checked quarterly. Corrections are published with a date rather than made silently.