SLPOtherClinical linguistics & phonetics2022

Comparison of auto-contouring and hand-contouring of ultrasound images of the tongue surface.

Kevin D Roon, Wei-Rong Chen, Rion Iwasaki and 5 others

PMID 34974782

WHAT IT FOUND

For ultrasound tongue contouring, hand tracing is consistent.

SLURP and EdgeTrak are closest to hand tracing, but all algorithms are less reliable at tongue tip and need checking.

Key findings

01Hand-placed tongue contours were very consistent, with a mean absolute difference of 0.51 mm between measurers at coextensive points.

02SLURP was the only algorithm whose mean coextensive difference was below the 1 mm threshold, while EdgeTrak, AAA2, and AAA2NS were slightly above and EPCS and AAA1 were well above.

03No algorithm met all thresholds, and all algorithms were less reliable than human measurers at the anterior tongue end.

STILL TO COME

How it was doneWhat they foundWhat it means for SLPs

Read the rest of this summary

You get three full summaries a month, free, and we do not ask for a card. Search, the TL;DRs and your library stay unlimited either way.

Already have one?

What it does not show

The study used only four speakers and 40 short clips, so it does not show how the methods perform on many clients or long recordings. The clips were about 748 ms long, much shorter than many real clinical or research recordings, so the authors could not tell whether some errors would grow, plateau, or correct with more time. The tolerance thresholds were based on human measurer differences plus an arbitrary margin, so passing or failing a threshold is not a fixed clinical accuracy standard. The hand contours were used as the standard, but there was no independent ground truth such as a calibrated concurrent imaging signal. All data came from one ultrasound system with fixed settings, and harder-to-see images may have made differences larger. The AAA methods were run on video files rather than original Cine Loop data, so their performance on Cine Loop data may be different. SLURP output changed slightly each run, although the authors judged that this did not meaningfully affect results.

Declared interests

The article text says the authors report no conflicts of interest. The supplied metadata indicates NIH extramural support, but no sponsor role is stated.

The easy way to misread this

Do not conclude that automatic tongue contouring is accurate enough for all speech sounds without checking. The study found that no algorithm met all thresholds, and all algorithms were much less reliable at the anterior tongue-tip end than human measurers.

Read it on PubMed →