Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “Sound Spectrography”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,135 records · Page 63Linked to original sources

Ventricular fold vibration in voice production: a high-speed imaging study with kymographic, acoustic and perceptual analyses of a voice patient and a vocally healthy subject.

Co-vibrations of the ventricular folds are a common finding in the clinical setting. It is not always obvious how much of the perceived voice change can be attributed to the presence of such vibrations. The aim of the present study was to describe laryngeal vibrations as observed by high-speed imaging in cases where ventricular fold vibrations had been observed. The findings at kymographic display of the recordings were correlated to perceptual measures and spectrographic observations. Two subjects, a 65-year-old man with chronic laryngitis and one vocally healthy man, were examined during pressed and breathy sustained phonation. Perceived roughness in the voice quality correlated to irregularities in true vocal fold vibrations as well as to irregular ventricular fold vibrations with large amplitude combined with sufficient closure. In none of the recorded sections did ventricular fold vibrations occur without simultaneous true vocal fold vibrations. Regular vibrations of the ventricular folds of the same frequency as those of the true vocal folds and with a reciprocal pattern did not contribute to any roughness in the perceived voice.

Adult↗

Measuring the tuning accuracy of thousands singing in unison: an English Premier Football League table of fans' singing tunefulness.

Tunefulness in singing is well understood in the context of solo stage performance, singing in small groups and singing in choirs, with or without accompaniment, and it can be readily measured under laboratory conditions. When thousands of people are singing outside in support of their football team, however, the singing is impromptu; there is no conductor, no starting note, and generally no accompaniment. This paper describes the measurement of the tunefulness of the singing of fans of the twenty clubs in the 2001-2002 English Premier League. The technique adopted is unusual in that it makes direct reference to the formal definition of pitch as a subjective phenomenon. The results are presented in the form of a 2001-2002 English Premier League football fans singing league table.

Acoustics↗

Does the acoustic waveform mirror the voice?

Over recent decades, much effort has been invested in the search for acoustic correlates of vocal function and dysfunction. The convenience of non-invasive voice measurements has been a major incentive for this effort. The acoustic signal is a rich but also very diversified source of information. Computer literacy and technical curiosity in the voice care and voice performance communities are now higher than ever, and tools for voice analysis are proliferating. On such a busy scene, a review may be useful of some basic principles for what we can and cannot hope to determine from non-invasive acoustic analysis. One way of doing this is to consider communication by voice as though it were engineered, with layered protocols. This results in a scheme for systematizing the many sources of variation that are present in the acoustic signal, that can complement other strategies for extracting information.

Humans↗

Speech prosody, voice quality and personality.

The 'individuation' of oral language is still largely an unknown territory, especially with respect to the individual use of speech prosody (rhythm, intonation and intensity). An initial study (Zellner Keller, 2004) has suggested an interesting relationship between speakers' prosodic styles and their personality profiles, as perceived by listeners. In this study, it is shown for 34 male speakers that distributions of F0 (for the intonation) and dB (for the intensity) give interesting information. F0 distributions among speakers are not similar, whether the speech task is the same or not, but intensity distributions are nearly perfectly superposed on each other. Individual intonational styles in relationship with personality styles are then examined in two tasks: reading texts and spontaneous speech.

Humans↗

Voice after supracricoid laryngectomy: subjective, objective and self-assessment data.

Supracricoid laryngectomy (SCL) is an efficient surgical procedure for the treatment of selected laryngeal carcinoma, presently being performed not only in Europe but also in North America. The functional goals of the technique are voice and swallowing without a permanent tracheostoma. Perceptual and acoustic voice characteristics after SCL have been reported by different authors, but self-assessment data together with subjective and objective data have only been reported for a small number of subjects. Twenty male subjects, with a mean age of 71 years (range: 51-82 years) who underwent a SCL at least one year before our observation, were included in the study. Each subject underwent a flexible laryngoscopy and his voice was perceptually rated using the GRBAS scale. Objective examination included: maximum phonation time (MPT), voice spectrograms and syllable diadochokinesis on a single breath. Finally, each subject assessed his own voice using the Voice Handicap Index (VHI). The mean values of the GRBAS scale were respectively 2.4, 2.6, 2.4, 0.8, 0.5, 0.8. Mean MPT was 7.5 s, while for voice spectrograms the mean value of the Yanagihara scale was 3.7. Mean syllable diadochokinesis appeared as 3.3 syllables/s. Mean value of the VHI was 29.9. Subjective and objective data show a severely dysphonic voice after SCL; self-assessment data, on the contrary, reveal only moderate functional and emotional consequences. While perceptual, aerodynamic and acoustic data are in line with previous reports, self-assessment data were less severe in our subjects compared to what appears in the literature. It is concluded that self-assessment explores a different dimension of the patient's voice and that even if a severe dysphonia is present the consequences on everyday oral communication are only moderate.

Aged↗

An exploratory baseline study of boy chorister vocal behaviour and development in an intensive professional context.

Currently, there is no existing published empirical longitudinal data on the singing behaviours and development of choristers who perform in UK cathedrals and major chapels. Longitudinal group data is needed to provide a baseline against which individual chorister development can be mapped. The choristers perform to a professional standard on a daily basis, usually with linked rehearsals, whilst also following a full school curriculum. The impact of this intensive schedule in relation to current vocal behaviour, health and future development requires investigation. Furthermore, it is also necessary to understand the relationship between the requirements of chorister singing behaviour and adolescent voice change. The paper will report the initial findings of a new longitudinal chorister study, based in one of London's cathedrals. Singing and vocal behaviours are being profiled on a six-monthly basis using data from a specially designed acoustic and behavioural instrument. The information obtained will enable us to understand better the effects of such training and performance on underlying vocal behaviour and vocal health. The findings will also have implications for singing teachers and choral directors in relation to particular methods of vocal education and rehearsal.

Adolescent↗

Objective analysis of the singing voice as a training aid.

A new tool for robust tracking of fundamental frequency is proposed, along with an objective measure of main singing voice parameters, such as vibrato rate, vibrato extent, and vocal intonation. High-resolution Power Spectral Density estimation is implemented, based on AutoRegressive models of suitable order, allowing reliable formant tracking also in vocalizations characterized by highly varying values. The proposed techniques are applied to about 1000 vocalizations, coming from both professional and non-professional singers, and show better performance as compared to classical Fourier-based approaches. If properly implemented, and with a user-friendly interface, the new tool would allow real-time analysis of singing voice. Hence, it could be of help in giving non-professional singers and singing teachers reliable measures of possible improvements during and after training.

Acoustics↗

Health and voice quality in smokers: an exploratory investigation.

Thirty-two adults (20 smokers and 12 non-smokers) were examined to determine the effects of cigarette smoking on health (diseases and larynx histology), on fundamental frequency (regularity and jitter) and stress level. The examination consisted of nasovideostroboscopy analysis, history case, electrolaryngography assessment for different task performance and self-assessment of stress. Although not statistically different, results indicate that smokers in comparison to non-smokers show: 1) slightly more health problems; 2) histological larynx changes; 3) a higher level of stress; 4) a lower mean F0 for all speech tasks. A statistically significant difference for percentage jitter was found according to voice status for: 1) vowel [a] (the group with voice problems show higher levels of jitter than both groups without voice problems); 2) vowel [i] (the smokers with voice problems show higher levels of jitter than the smokers without voice problems).

Adult↗

Mirroring the voice from Garcia to the present day: some insights into singing voice registers.

Starting from Garcia's definition, the historical evolution of the notion of vocal registers from then until now is considered. Even though much research has been carried out on vocal registers since then, the notion of registers is still confused in the singing voice community, and defined in many different ways. While some authors consider a vocal register as a totally laryngeal event, others define it in terms of overall voice quality similarities. This confusion is reflected in the multiplicity of labellings, and it lies in the difficulty of identifying and specifying the mechanisms distinguished by these terms. The concept of laryngeal mechanism is then introduced, on the basis of laryngeal transition phenomena detected by means of electroglottography. It helps to specify at least the laryngeal nature of a given singing voice register. On this basis, the main physiological, acoustic, and perceptual characteristics of the most common singing voice registers are surveyed.

History, 19th Century↗

Discussant response to 'Does the acoustic waveform mirror the voice?'.

The acoustic waveform that reaches the two ears of a listener can convey the intended message. Over the telephone, this waveform is the only source of information from the speaker, since the listener is out of visual contact; indeed, in this situation the acoustic waveform itself is restricted in its frequency content. Whilst the listener can infer much about the speaker from the acoustic waveform, including the speaker's age, gender, nationality, dialect, and emotional state, the issue under consideration here is the extent to which quantitative analysis of the acoustic waveform can provide useful information about a speaker's voice. This paper was presented as an Invited Discussant Response, at the Pan-European Voice Conference (PEVOC6) in London, to the question posed in Ternström's Invited Keynote Lecture 1: Does the acoustic waveform mirror the voice?

Humans↗

The intelligibility of tracheoesophageal speech, with an emphasis on the voiced-voiceless distinction.

Total laryngectomy has far-reaching effects on vocal tract anatomy and physiology. The preferred method for restoring postlaryngectomy oral communication is prosthetic tracheoesophageal (TE) speech, which like laryngeal speech is pulmonary driven. TE speech quality is better than esophageal or electrolarynx speech quality, but still very deviant from laryngeal speech. For a better understanding of neoglottis physiology and for improving rehabilitation results, study of TE speech intelligibility remains important. Methods used were perceptual evaluation, acoustic analyses, and digital high-speed imaging. First results show large variations between speakers and especially difficulty in producing voiced-voiceless distinction. This paper discusses first results of our experiment.

Adult↗

Acoustic characteristics of crying in infantile laryngomalacia.

The purpose of this study was to add to the extant data base of acoustic cry studies by profiling the condition of laryngomalacia. We hypothesized that the acoustic characteristics of crying produced by an infant with laryngomalacia would differ compared to previously reported cry data for normal infants. An entire episode of crying was audio recorded and acoustically analyzed for the occurrence of expiratory and inspiratory cry segments, as well as the long-time average spectral (LTAS) characteristics. Results obtained for the infant were found to be considerably different from what has been previously reported for normal infants. The overall duration of the infant's crying episode was longer, with proportionately fewer expiratory phonations and more inspiratory phonations compared to normal infants. The LTAS results were reflective of aperiodic components in the glottal source spectrum. Collectively, the infant's unusual crying aspects were not limited solely to those acoustic features resulting from a prolapse of supraglottic soft tissue, and therefore provide new insight into the vocal fold vibratory behavior characterizing infantile laryngomalacia.

Acoustics↗

Effects of phase changes in low-numbered harmonics on the internal representation of complex sounds.

A series of experiments investigated the effect of phase changes in low-numbered single harmonics in target sounds that were either synthesized steady-state vowels fo periodic signals having only a single formant. A matching procedure was sued in which subjects selected a sound along a continuum differing in first formant frequency in order to get the best match with the target sound; perceptual effects of the phase manipulations in the target were detected as a change in the matched first formant frequency. Stimuli had to contain at least three harmonics to produce the effect, but id did not require a particular starting phase of the components. A suppression phenomenon is discussed, in which phase changes alter the phase-locking characteristics of auditory fibres tuned to low-numbered harmonics.

Adult↗

Spectral weights and the profile bowl.

In these experiments, listeners detected changes in the shape of a complex spectrum that varied in overall level. With multicomponent complexes, a typical finding is that the listeners were more sensitive to changes made in the middle components of the spectrum than to changes made at either edge. We used a technique developed by Berg (1989) to estimate the weight listeners attached to the different components of the spectrum in making these judgements. For reasons not understood, the pattern of spectral weights was nearly optimum for a change made in the middle component of the spectrum and much poorer when the change occurred at either edge.

Attention↗

Perception of spectral changes in multi-tone complexes.

Amplitude changes of the spectral components of a complex tone, relative to each other, are usually well perceived, even if the over-all intensity is kept fixed. Three experiments are reported: Experiment 1 dealt with the detectability of amplitude changes in two-tone complexes of fixed frequencies. Experiment 2 examined detection of slope changes in ramp-shaped spectral envelopes of two- and three-tone complexes as a function of spectral spacing. As a control experiment for some conditions a roving intensity level was used. Experiment 3 investigated the detectability of changes in the spectral slope of multi-tone complexes as a function of the number of components. The results of the experiments show that detection of spectral changes in a sound is strongly dependent on the frequency spacing of the components. It is concluded that the auditory system is capable of comparing the relative energy distributions over different critical bands. Within a critical band there exists an optimum frequency separation with respect to the detection of relative amplitude change.

Adult↗

Formant transition duration and amplitude rise time as cues to the stop/glide distinction.

There is some disagreement in the literature about the relative contribution of formant transition duration and amplitude rise time in signalling the contrast between stops and glides. In this study, listeners identified sets of /ba/ and /wa/ stimuli in which transition duration and rise time varied orthogonally. Both variables affected labelling performance in the expected direction (i.e. the proportion of /b/ responses increased with shorter transition durations and shorter rise times). However, transition duration served as the primary cue to the stop/glide distinction, whereas rise time played a secondary, contrast-enhancing role. A qualitatively similar pattern of results was obtained when listeners made abrupt-onset/gradual-onset judgements of single sine-wave stimuli that modelled the rise times, frequency trajectories, and durations of the first formant in the /ba/-/wa/ stimuli. The similarities between the speech and non-speech conditions suggest that significant auditory commonalities underlie performance in the two cases.

Adult↗

Effects of impulse noise on transiently evoked otoacoustic emission in soldiers.

The aim of the study was to assess the effects of exposure to impulse noise on TEOAE, as compared to PTA. The study comprised 92 soldiers, subjected to impulse noise during military service. The control group consisted of secondary school students, not exposed to noise. Extended high frequency PTA, and TEOAE were recorded before and after one year of military service. The total level of noise and spectrum analysis were performed for all kinds of weapons, separately. The highest levels of noise for weapons were related to frequencies from 1.6-16 kHz. After military service significant deterioration of hearing was observed on average by 6 dB exclusively at the frequencies of 10 and 12 kHz. TEOAE reduction was registered predominantly at frequencies of 2, 3 and 4 kHz, with the greatest decrease at 2 kHz (p <0.02). The control group did not show any significant audiometric changes as well as TEOAE during the time of experiment.

Acoustic Stimulation↗

Analysis of counted behaviors in a single-subject design: modeling of hearing-aid intervention in hearing-impaired patients with Alzheimer's disease.

Clinical procedures related to patients with Alzheimer's Disease (AD) largely fail to address the patient's hearing. Given the challenges of this population, unconventional indicators of treatment efficacy may be required. Palmer et al (1999) reported on caregiver-tracked behaviors as outcome measures for hearing aid intervention. Using these data, hearing aid use and subsequent behavior was modeled as a first-order dynamic system, characterized by responses following an exponential time course. The results of such modeling suggest predictable outcomes of hearing aid intervention, or at least useful parameters of quantification (e.g. time-constant and steady-state response), permitting critical assessment of effects of intervention on negative behaviors versus hearing aid use, comparisons among behaviors, and/or comparisons of hearing-aid-use patterns and behavior counts among patients. Use in this and other difficult-to-test populations warrant further study to evaluate clinical efficacy of the analysis described.

Aged↗