Performance in an Audiovisual Selective Attention Task Using Speech-Like Stimuli Depends on the Talker Identities, But Not Temporal Coherence.
Madeline S Cappelloni, Vincent S Mateo, Ross K Maddox
PMID 37847849WHAT IT FOUND
Seeing the talker's face helped listeners ignore background noise better than matching lip movements did.
The benefit came from recognizing who was speaking, which reduced errors in identifying the wrong voice. This suggests face identity aids attention more than visual timing.
Key findings
01Matching the video talker to the target voice significantly improved performance, reducing false alarms.
02Temporal coherence (matching lip movement timing) provided no measurable benefit to performance in this task.
03Vowel identity matching had no significant effect on task performance.
STILL TO COME
How it was doneWhat they foundWhat it means for SLPs
Read the rest of this summary
You get three full summaries a month, free, and we do not ask for a card. Search, the TL;DRs and your library stay unlimited either way.
What it does not show
The stimuli were artificial speech-like vowels, not natural conversational speech, which may limit generalizability to real-world listening. Only two talkers (one male, one female) were used, so it is unclear if the benefit of talker identity is due to gender differences or individual voice/face characteristics. The task required detecting pitch changes in a controlled setting, which may not reflect the complexity of understanding speech in everyday noise.
Declared interests
The authors declared no potential conflicts of interest.
The easy way to misread this
Do not assume that matching lip movements to speech sounds is unimportant for all patients. This study used artificial vowel stimuli where segregation was relatively easy; in more complex or natural speech scenarios, temporal coherence may still play a role in binding.