Aug 2026· Journal of the Acoustical Society of America· Vol 159, pp. A320-A320· 0 citations
Abstract
Perception of speech sounds is shaped by the context of the earlier sounds, including speaking rate. Often, acoustic differences between earlier and later sounds are perceptually magnified. In temporal contrast effects (TCEs), a fast-rate precursor sentence can cause the following target word to sound longer (e.g., longer-VOT “tier”), and a slow-rate precursor sentence can cause the target word to sound shorter (e.g., shorter-VOT “deer”). The novel contribution of this study is the exploration of TCEs across different context durations to examine their consistency on a granular level. On each trial, listeners heard one of three contexts (“a”, “the word”, “this time I want you to click on the word”) spoken at a fast or slow rate, then a target word to be identified as “deer” or “tier”. TCEs were observed at all context durations but were surprisingly not consistent in magnitude. While speaking rate forms an important context for speech sound recognition, the degree of its influence may not be reliable across different timescales. This reveals important bounds on how context effects relate to one another. [Work supported by NIDCD.]
The temporal structure of speech has traditionally been characterized by the rhythmicity of its canonical linguistic units (phonemes, syllables, words), each summarized by a mean occurrence rate. While valid, this view overlooks whether speech carries a finer, context-dependent and probabilistic temporal structure that...
Laure Deyna, Philippe Albouy, Agnès Trébuchon et al.· bioRxiv· 1 citation
Every syllable and every phoneme in natural speech has a different duration. This variability is conventionally treated as jitter, noise the brain must overcome to recover an underlying regularity. Here we show that the cortex encodes it as information. We recorded magnetoencephalography from 25 native listeners of Spa...
N. Molinaro, Jose Pérez-Navarro· bioRxiv· 0 citations
Predicting when an event will occur is key to perceiving and evaluating dynamic sensory inputs, such as speech. The utility of temporal predictions hinges on perceptual timing precision (PTP), the sensitivity to deviations between actual and predicted sound onset times. Past work suggests that PTP is primarily determin...
Charlotte M. Mock, Leon Ilge, Yulia Oganian· bioRxiv· 0 citations
In cocktail party-type environments, listeners must segregate simultaneous acoustic sources and select one for processing. Behavioral studies have established that both low-level voice differences (e.g., pitch) and high-level linguistic differences (e.g., word content) aid these processes. Electroencephalography (EEG)...
Sahil Luthra, Eric Parker, Marysia Brown et al.· bioRxiv· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.