Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “Sound Spectrography”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 865 records · Page 48Linked to original sources

Speech interactions with linguistic, cognitive, and visuomotor tasks.

Lip movements were examined across several repetitive speaking conditions (speech alone and speaking concurrently with a linguistic, cognitive, or visuomotor challenge task) in 20 young adults. Performance in these nonspeech activities was also compared between isolated tasks and concurrent speech conditions. Linguistic challenges resulted in increased spatiotemporal variability of lip displacement across repetitions. Motor challenges led to more rapid speech with smaller lip displacement. These qualitatively different changes suggest that different aspects of attention are required for linguistic versus manual visuomotor activity. Vocal intensity increased for all concurrent task conditions compared with speech alone, suggesting increased effort compared to the control condition. Scores for linguistic performance decreased when utterance repetition occurred concurrently with the syntactic challenge. These findings reveal that speech motor activity can influence linguistic performance as well as be influenced by it. Although these data come from healthy speakers, they suggest that clinicians working with disordered speakers should not overlook the potential interactions among the demands of language formulation, cognitive activity, and speech motor performance.

Adult↗

Measurement of vocal fold collision forces during phonation: methods and preliminary data.

Forces applied to vocal fold tissue as the vocal folds collide may cause tissue injury that manifests as benign organic lesions. A novel method for measuring this quantity in humans in vivo uses a low-profile force sensor that extends along the length and depth of the glottis. Sensor design facilitates its placement and stabilization so that phonation can be initiated and maintained while it is in place, with minimal interference in vocal fold vibration. In 2 individuals with 1 vibrating vocal fold and 1 nonvibrating vocal fold, peak collision force correlates more strongly with voice intensity than pitch. Vocal fold collision forces in 1 individual with 2 vibrating vocal folds are of the same order of magnitude as in previous studies. Correlations among peak collision force, voice intensity, and pitch were indeterminate in this participant because of the small number of data points. Sensor modifications are proposed so that it can be used to reliably estimate collision force in individuals with 2 vibrating vocal folds and with changing vocal tract conformations.

Biomechanical Phenomena↗

Adaptation of a Pocket PC for use as a wearable voice dosimeter.

This article deals with the adaptation of a commercially available Pocket PC for use as a voice dosimeter, a wearable device that measures the vocal dose of teachers or other individuals on the job, at home, and elsewhere during the course of an entire day. An engineering approach for designing a voice dosimeter is described, and design data are presented. Technical issues include transducer selection, dynamic range, frequency response, memory requirements, power requirements, attachment, cables, connections, and data collection. Advantages and disadvantages of the design are discussed.

Electric Stimulation↗

Multiple looks in speech sound discrimination in adults.

N. F. Viemeister and G. H. Wakefield's (1991) multiple looks hypothesis is a theoretical approach from the psychoacoustic literature that has promise for bridging the gap between results from speech perception research and results from psychoacoustic research. This hypothesis accounts for sensory detection data and predicts that if the "looks" at a stimulus are independent and information is combined optimally, sensitivity should increase for 2 pulses relative to 1 pulse. Specifically, d' (a bias-free measure of sensitivity) for 2 pulses should be larger than d' for 1 pulse. One speech discrimination paradigm that presents stimuli with multiple presentations is the change/no-change procedure. On a change trial, the standard and comparison stimuli differ; on a no-change trial, they are the same. Normal-hearing adults were tested using the change/no-change procedure with 3 consonant-vowel minimal pairs in combinations of 1, 2, and 4 repetitions of standard and comparison stimuli at various signal-to-noise ratios. If multiple looks extend to this procedure, performance should increase with higher repetition numbers. Performance increased with more presentations of the speech contrasts tested. The multiple looks hypothesis predicted performance better at low repetition numbers when performance was near d' values of 1.0 than at higher repetition numbers and higher performance levels.

Adult↗

Control of voice-onset time in the absence of hearing: a review.

The relation between partial or absent hearing and control of the voicing contrast has long been of interest to investigators, in part because speakers who are born deaf characteristically have great difficulty mastering the contrast and in part for the light it can cast on the role of hearing in the acquisition and maintenance of phonological contrasts in general. One of the phonetic characteristics that distinguish voiced from voiceless plosives in English (p/b, t/d, k/g) is voice onset time (VOT): the interval from plosive release to the onset of voicing of the following vowel. This article first reviews research on VOT anomalies in the speech production of prelingually and postlingually deaf speakers. Then it turns to studies of the mechanisms in speech breathing, phonation and articulation that underlie those anomalies. In both populations of speakers, there is a tendency for the difference between voiced and voiceless VOT to be reduced, to the point for many speakers that there is in effect a substitution of the voiced for the voiceless cognate. The separation of the cognate VOTs can be enhanced when some hearing is restored with a cochlear implant. Both populations also present anomalies in speech breathing that can hinder the development of intraoral pressures and transglottal pressure drops that are required for the production of the VOT contrast. Its successful management further requires critical timing among phonatory and articulatory gestures, most of which are not visible, rendering the VOT contrast a particular challenge in the absence of hearing.

Articulation Disorders↗

Parametric quantitative acoustic analysis of conversation produced by speakers with dysarthria and healthy speakers.

PURPOSE: This study's main purpose was to (a) identify acoustic signatures of hypokinetic dysarthria (HKD) that are robust to phonetic variation in conversational speech and (b) determine specific characteristics of the variability associated with HKD. METHOD: Twenty healthy control (HC) participants and 20 participants with HKD associated with idiopathic Parkinson's disease (PD) repeated 3 isolated sentences (controlled phonetic content) and 2 min of conversational speech (phonetic content treated as a random variable). A MATLAB-based program automatically calculated measures of contrastivity: speech-pause ratio, intensity variation, median and maximum formant slope, formant range, change in the upper and lower spectral envelope, and range of the spectral envelope. t tests were used to identify which measures were sensitive to HKD and which measures differed by task. Discriminant analysis was used to identify the combination of measures that best predicted HKD, and this analysis was then used as a general measure of contrastivity (Contrastivity Index). Differential effects of HKD on maximum and typical contrastivity levels were tested with interaction of maximum, minimum, and median observations of individual speakers and with pairwise comparisons of skewness and kurtosis of the contrastivity index distributions. RESULTS: Group differences were detected with pairwise comparisons with t tests in 8 of the 9 measures. Percentage pause time and spectral range were identified as the most specific (95%) and accurate (95%) differentiators of HKD and HC conversational speech. Sentence repetition elicited significantly higher levels of contrastivity than conversational speech in both HC and HKD speakers. Maximum and minimum contrastivities were significantly lower in HKD speech, but there was no evidence that HKD affects maximum contrastivity levels more than median contrastivity levels. The HKD speakers' contrastivity distributions were significantly more skewed to lower levels of production. CONCLUSION: HKD can be consistently distinguished from HC speech in both sentence repetition and conversational speech on the basis of intensity variation and spectral range. Although speakers with HKD were effectively able to produce higher contrastivity levels in sentence repetition tasks, they habitually performed closer to the lower end of their production ranges.

Adult↗

Effect of facemask use on respiratory patterns of women in speech and singing.

PURPOSE: Research into respiratory behavior during singing and speech makes extensive use of standard respiratory and vented pneumotachograph facemasks. This study investigated whether the use of such facemasks would affect respiratory behavior in terms of lung volume (excursion, at initiation and at termination) or duration (of inspiration and of expiration) during speech or singing. METHOD: The respiratory patterns of 6 females were recorded using uniaxial surface magnetometry during 4 tasks: quiet breathing, a /pa/ syllabic train, reading ("The Rainbow Passage"), and singing a Christmas carol ("Silent night"). Each task was performed in 4 facemask conditions: wearing no facemask, wearing a facemask rim only, wearing a standard respiratory facemask, and wearing a vented pneumotachograph facemask. RESULTS: No significant effect was found for any of the facemask conditions on lung volume or duration measures during any tasks. CONCLUSION: The results confirm earlier studies that the vented pneumotachograph facemask does not affect breathing behavior in speech research studies and extends the finding to the study of breathing behavior in singing and to the use of a standard respiratory facemask.

Adult↗

Voice training and therapy with a semi-occluded vocal tract: rationale and scientific underpinnings.

PURPOSE: Voice therapy with a semi-occluded vocal tract has a long history. The use of lip trills, tongue trills, bilabial fricatives, humming, and phonation into tubes or straws has been hailed by clinicians, singing teachers, and voice coaches as efficacious for training and rehabilitation. Little has been done, however, to provide the scientific underpinnings. The purpose of the study was to investigate the underlying physical principles behind the training and therapy approaches that use semi-occluded vocal tract shapes. METHOD: Computer simulation, with a self-oscillating vocal fold model and a 44 section vocal tract, was used to elucidate source-filter interactions for lip and epilarynx tube semi-occlusions. RESULTS: A semi-occlusion in the front of the vocal tract (at the lips) heightens source-tract interaction by raising the mean supraglottal and intraglottal pressures. Impedance matching by vocal fold adduction and epilarynx tube narrowing can then make the voice more efficient and more economic (in terms of tissue collision). CONCLUSION: The efficacious effects of a lip semi-occlusion can also be realized for nonoccluded vocal tracts by a combination of vocal fold adduction and epilarynx tube adjustments. It is reasoned that therapy approaches are designed to match the glottal impedance to the input impedance of the vocal tract.

Computer Simulation↗

The relationship of vocal loudness manipulation to prosodic F0 and durational variables in healthy adults.

This investigation was motivated by observations that when persons with dysarthria increase loudness their speech improves. Some studies have indicated that this improvement may be related to an increase of prosodic variation. Studies have reported an increase of fundamental frequency (F0) variation with increased loudness, but there has been no examination of the relation of loudness manipulation to specific prosodic variables that are known to aid a listener in parsing out meaningful information. This study examined the relation of vocal loudness production to selected acoustic variables known to inform listeners of phrase and sentence boundaries: specifically, F0 declination and final-word lengthening. Ten young, healthy women were audio-recorded while they read aloud a paragraph at what each considered normal loudness, twice-normal loudness, and half-normal loudness. Results showed that there was a statistically significant increase of F0 declination, brought about by a higher resetting of F0 at the beginning of a sentence and an increase of final-word lengthening from the half-normal loudness condition to the twice-normal loudness condition. These results suggest that when some persons with dysarthria increase loudness, variables related to prosody may change, which in turn contributes to improvement in communicative effectiveness. However, until this procedure is tested with individuals who have dysarthria, it is uncertain whether a similar effect would be observed.

Adolescent↗

Tennessee Test of Rhythm and Intonation Patterns.

The Tennessee Test of Rhythm and Intonation Patterns (T-TRIP) is a three-part suprasegmental test with 25 test items. The test items consisted of the nonsense syllable (ma) that was spoken and recorded with different rhythm and intonation patterns. Ten three-year-olds and 10 five-year-olds imitated the pattern they heard. The five-year-olds scored significantly better than the three-year-olds. The T-TRIP appears sensitive to differences between groups of different ages.

Age Factors↗

The cry characteristics of an infant who died of the sudden infant death syndrome.

Fourteen cries of a four day old infant who subsequently died suddenly of unexplained causes were analyzed on nine acoustic characteristics including fo, duration, formant frequencies and sound pressure level. In comparison to a group of newborn controls, the Sudden Infant Death Syndrome (SIDS) victim's cries exhibited a lower fo, longer duration, lower formant frequencies and greater sound pressure level throughout the spectrum. Cry duration and sound pressure levels, however, deviated in excess on one standard deviation from the mean of the other newborns. Similar findings resulted when the SIDS infant was compared to a group of full term infants who were siblings of SIDS victims, although the magnitude of the differences was slightly less especially with respect to sound pressure level. Measurement of selected acoustic variables in a newborn's cry may be of value in our understanding of SIDS and for identifying infants at risk.

Acoustics↗

The speech spectrographic display: interpretation of visual patterns by hearing-impaired adults.

If visual speech training aids are to be used effectively, it is important to assess whether hearing-impaired speakers can accurately interpret visual patterns and arrive at correct conclusions concerning the accuracy of speech production. In this investigation with the Speech Spectrographic Display (SSD), a pattern interpretation task was given to 10 hearing-impaired adults. Subjects viewed selected SSD patterns from hearing-impaired speakers, evaluated the accuracy of speech production, and identified the SSD visual features that were used in the evaluation. In general, results showed that subjects could use SSD patterns to evaluate speech production. For those pattern interpretation errors that occurred most were related either to phonetic/orthographic confusions or to misconceptions concerning production of speech.

Adult↗

Compensatory articulation patterns of a surgically treated oral cancer patient.

Acoustic, physiologic, and perceptual data on articulation of bilabial and alveolar stop consonants in CVC words spoken by a subject with a 20% glossectomy, 2/3-mandibulectomy and radical neck dissection, whose surgical closure was accomplished by suturing the tongue to the floor of the mouth and buccal mucosa are provided. Compensatory articulation patterns were identified from tracings of vocal tract shapes obtained by videofluoroscopy. Acoustic analysis was accomplished by means of computer-aided examination of the speech waveform. Compensatory postures for articulation of the stop consonants were characterized by varying degrees of labial protrusion and retraction, dependent on the vowel context. Distinction between the bilabial and tip-alveolar consonants was evidenced in lingual-velar closure vs. lingual-velar-palatal closure. These compensatory patterns are discussed in relation to the acoustic properties of the corresponding speech sounds obtained from synthetic speech samples and in relation to perceptual data obtained from trained listeners.

Glossectomy↗

Phonetic disintegration in a five-year-old following sudden hearing loss.

The speech of a five-year-old boy who suffered a profound hearing loss following meningitis was sampled at two-week intervals for nine months. Speech samples were subjected to phonetic transcription, spectrographic analysis, and intelligibility testing. Immediately post-trauma, the child displayed slightly slower, F0 elevated, acoustically intense speech in which phonemic distortion and syllabification of consonants occurred occasionally; single word intelligibility was depressed below normal between 20-30%. By the 18th week, a sudden decline in intelligibility, increasing monotony of pitch, and a pattern of strongly emphatic, prolonged, aspirated, syllabified, and increasingly distorted consonants were manifest. At year's end, the child's speech bore some resemblance to the speech of the deaf in terms of suprasegmentals, intonation, and intelligibility, but differed because the child rarely, if ever deleted speech sounds or diphthongized vowels strongly. It is speculated that phonetic processes such as diphthongization, syllabification, and prolonged duration may be strategies for enhancing feedback during speech.

Child, Preschool↗

Perception and production of misarticulated (r).

Twelve children who consistently misarticulated consonant [r] and five children who correctly articulated [r] were recorded while repeating sentences which differed only in a single (r)-(w) contrast. All (r) and (w) productions were spectrographically analyzed. Error productions were judged for their similarity to [w]. Each child identified all of the recorded sentences via a picture-pointing task. Misarticulated [r] was identified as (w) at above chance levels only by the children who did not misarticulate [r]. The subject groups did not differ in their perception of correctly articulated (r) and (w) phones. Children whose misarticulated [r] phones were judged to be (w)-like were most likely to misperceive their own productions of (r). Children whose misarticulated [r] productions were characterized by higher second formant frequencies were better able to identify their productions of (r). Results suggest that a subpopulation of children who misarticulate [r] may mark it acoustically in a nonstandard manner.

Articulation Disorders↗