Skip to main content
AI Autopilot — +6% throughput and $781K annual profit gain in a 25-day controlled pilot. Brasil Mineral #441 (Jul 2024).See Pilot Results
Brainiall
  • Products
  • Sectors
  • Demo
  • Pilot Results
  • Why Us
Sign InTry Live Demo
Brainiall
AI that runs the real economy. Mining. Industrial. Voice. Bootstrapped since 2019.
Products
  • AI Autopilot
  • Specialist AIs
Sectors
  • Industrial
  • Mining
  • Energy
  • Technology
Company
  • About
  • Why Choose Us
  • Pilot Results
  • Articles
  • Changelog
  • Contact Us
Resources
  • Pricing
  • Developer APIs
  • Docs
  • Integrations
  • Compare
  • Trust Center
  • DPA
  • MSA
  • Status

Stay updated

Specialist AI insights, delivered weekly.

© 2026 Brainiall, Inc. All rights reserved.
Privacy PolicyTerms of Use
Home/For language teachers

For language teachers

How do you measure a student’s pronunciation without sitting through every take?

A listener is a rater. Two trained raters, scoring the same phones, agree with each other at Phone PCC 0.555. The engine behind the free checker sits at Phone PCC 0.682 against that same human consensus. That gap is why a number can replace “it sounded fine.”

0–39Significant error
40–59Noticeable
60–79Acceptable
80–100Native-like
Phoneme bands from the scoring guide. Every threshold links to the scale.

Forced alignment against a phrase, not a vibe

The engine does not grade “accent” and it does not treat free transcription as a mark. You send audio plus the exact sentence the student was asked to say. Forced alignment locks each phone to that reference. If they said a different sentence, that shows up as a mismatch — decodedTranscript versus the phrase — not as a quietly inflated overallScore.

That is the difference between a pronunciation meter and a speech-to-text box with a thumbs-up. The meter only exists because the reference text is known.

Four scores, four jobs

Every assessment returns the same four 0–100 layers: overallScore (confidence-weighted composite), sentenceScore (fluency, pacing, connected speech), per-word scores, and per-phoneme scores in IPA and ARPAbet. Words below 50 almost always hide one or more phonemes in 0–39 or 40–59.

You also get confidence (0 to 1) and, when the take is dirty, warnings. A high overallScore with confidence under 0.5 is not a high score. Throw the take out and re-record.

  • Same phrase for the whole class — otherwise you are not measuring the same thing.
  • Sort by words below 50, not by who “sounded better.”
  • Ignore any row whose confidence is under 0.5. That is the engine telling you the audio and the phrase do not match.

What this looks like in a class of twenty

Twenty files of the same sentence. You do not play twenty files. You read twenty word tapes. The students whose DH or IH collapsed are the ones who get a note tonight; the rest do not need a teacher’s ear on that homework. Phone PCC 0.682 versus 0.555 is the reason you can do that without pretending two human listeners would have agreed anyway.

Measure one take, not a theory

Type the phrase you actually assign. Record it. The free checker returns overallScore and a colour per word from the same engine.

Score a recording now