Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “Sound Spectrography”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,765 records · Page 98Linked to original sources

Objective evaluation of chamber-music halls in Europe and Japan.

The room acoustical parameters reverberation time, RT; early decay time, EDT; clarity, C80; time gravity, Tg; bass ratio, BR; strength, G; initial time delay gap, ITDG; interaural cross-correlation coefficient, IACC(E), the where binaural quality index BQI equals [1-IACC(E3)]; and stage support, ST1 were measured in 18 major chamber-music halls in Austria, Germany, the Netherlands, Czech Republic, Switzerland, and Japan, employing procedures in accordance with ISO 3382 (1997). In combination with the architectural data, the intrinsic objective parameters for the acoustics of chamber-music halls and their variation range were examined. The results of these studies reveal four pertinent orthogonal parameters: RT, G, ITDG, BQI. General design guidelines for a chamber-music hall are presented.

Acoustics↗

Sounds produced by Norwegian killer whales, Orcinus orca, during capture.

To date very little is still known about the acoustic behavior of Norwegian killer whales, in particular that of individual whales. In this study a unique opportunity was presented to document the sounds produced by five captured killer whales in the Vestfjord area, northern Norway. Individuals produced 14 discrete and 7 compound calls. Two call types were used both by individuals 16178 and 23365 suggesting that they may belong to the same pod. Comparisons with calls documented in Strager (1993) showed that none of the call types used by the captured individuals were present. The lack of these calls in the available literature suggests that call variability within individuals is likely to be large. This short note adds to our knowledge of the vocal repertoire of this population and demonstrates the need for further studies to provide behavioural context to these sounds.

Acoustics↗

Repetition patterns in Weddell seal (Leptonychotes weddellii) underwater multiple element calls.

Many vocalizations produced by Weddell seals (Leptonychotes weddellii) are made up of repeated individual distinct sounds (elements). Patterning of multiple element calls was examined during the breeding season at Casey and Davis, Antarctica. Element and interval durations were measured from 405 calls all > 3 elements in length. The duration of the calls (22+/-16.6 s) did not seem to vary with an increasing number of elements (F4,404=1.83,p = 0.122) because element and interval durations decreased as the number of elements within a call increased. Underwater vocalizations showed seven distinct timing patterns of increasing, decreasing, or constant element and interval durations throughout the calls. One call type occurred with six rhythm patterns, although the majority exhibited only two rhythms. Some call types also displayed steady frequency changes as they progressed. Weddell seal multiple element calls are rhythmically repeated and thus the durations of the elements and intervals within a call occur in a regular manner. Rhythmical repetition used during vocal communication likely enhances the probability of a call being detected and has important implications for the extent to which the seals can successfully transmit information over long distances and during times of high level background noise.

Acoustics↗

Amplification and spectral shifts of vocalizations inside burrows of the frog Eupsophus calcaratus (Leptodactylidae).

A variety of animals that communicate by sound emit signals from sites favoring their propagation, thereby increasing the range over which these sounds convey information. A different significance of calling sites has been reported for burrowing frogs Eupsophus emiliopugini from southern Chile: the cavities from which these frogs vocalize amplify conspecific vocalizations generated externally, thus providing a means to enhance the reception of neighbor's vocalizations in chorusing aggregations. In the current study the amplification of vocalizations of a related species, E. calcaratus, is investigated, to explore the extent of sound enhancement reported previously. Advertisement calls broadcast through a loudspeaker placed in the vicinity of a burrow, monitored with small microphones, are amplified by up to 18 dB inside cavities relative to outside. The fundamental resonant frequency of burrows, measured with broadcast noise and pure tones, ranges from 842 to 1836 Hz and is significantly correlated with the burrow's length. Burrows change the spectral envelope of incoming calls by increasing the amplitude of lower relative to higher harmonics. The call amplification effect inside burrows of E. calcaratus parallels the effect reported previously for E. emiliopugini, and indicates that the acoustic properties of calling sites may affect signal reception by burrowing animals.

Acoustics↗

Detection of random alterations to time-varying musical instrument spectra.

The time-varying spectra of eight musical instrument sounds were randomly altered by a time-invariant process to determine how detection of spectral alteration varies with degree of alteration, instrument, musical experience, and spectral variation. Sounds were resynthesized with centroids equalized to the original sounds, with frequencies harmonically flattened, and with average spectral error levels of 8%, 16%, 24%, 32%, and 48%. Listeners were asked to discriminate the randomly altered sounds from reference sounds resynthesized from the original data. For all eight instruments, discrimination was very good for the 32% and 48% error levels, moderate for the 16% and 24% error levels, and poor for the 8% error levels. When the error levels were 16%, 24%, and 32%, the scores of musically experienced listeners were found to be significantly better than the scores of listeners with no musical experience. Also, in this same error level range, discrimination was significantly affected by the instrument tested. For error levels of 16% and 24%, discrimination scores were significantly, but negatively correlated with measures of spectral incoherence and normalized centroid deviation on unaltered instrument spectra, suggesting that the presence of dynamic spectral variations tends to increase the difficulty of detecting spectral alterations. Correlation between discrimination and a measure of spectral irregularity was comparatively low.

Acoustic Stimulation↗

An acoustic description of the vowels of Northern and Southern Standard Dutch.

A database is presented of measurements of the fundamental frequency, the frequencies of the first three formants, and the duration of the 15 vowels of Standard Dutch as spoken in the Netherlands (Northern Standard Dutch) and in Belgium (Southern Standard Dutch). The speech material consisted of read monosyllabic utterances in a neutral consonantal context (i.e., /sVs/). Recordings were made for 20 female talkers and 20 male talkers, who were stratified for the factors age, gender, and region. Of the 40 talkers, 20 spoke Northern Standard Dutch and 20 spoke Southern Standard Dutch. The results indicated that the nine monophthongal Dutch vowels /a [see symbol in text] epsilon i I [see symbol in text] u y Y/ can be separated fairly well given their steady-state characteristics, while the long mid vowels /e o ø/ and three diphthongal vowels /epsilon I [see symbol in text]u oey/ also require information about their dynamic characteristics. The analysis of the formant values indicated that Northern Standard Dutch and Southern Standard Dutch differ little in the formant frequencies at steady-state for the nine monophthongal vowels. Larger differences between these two language varieties were found for the dynamic specifications of the three long mid vowels, and, to a lesser extent, of the three diphthongal vowels.

Adult↗

The effect of overlap-masking on binaural reverberant word intelligibility.

Reverberation interferes with the ability to understand speech in rooms. Overlap-masking explains this degradation by assuming reverberant phonemes endure in time and mask subsequent reverberant phonemes. Most listeners benefit from binaural listening when reverberation exists, indicating that the listener's binaural system processes the two channels to reduce the reverberation. This paper investigates the hypothesis that the binaural word intelligibility advantage found in reverberation is a result of binaural overlap-masking release with the reverberation acting as masking noise. The tests utilize phonetically balanced word lists (ANSI-S3.2 1989), that are presented diotically and binaurally with recorded reverberation and reverberation-like noise. A small room, 62 m3, reverberates the words. These are recorded using two microphones without additional noise sources. The reverberation-like noise is a modified form of these recordings and has a similar spectral content. It does not contain binaural localization cues due to a phase randomization procedure. Listening to the reverberant words binaurally improves the intelligibility by 6.0% over diotic listening. The binaural intelligibility advantage for reverberation-like noise is only 2.6%. This indicates that binaural overlap-masking release is insufficient to explain the entire binaural word intelligibility advantage in reverberation.

Acoustics↗

Geographic variation and acoustic structure of the underwater vocalization of harbor seal (Phoca vitulina) in Norway, Sweden and Scotland.

The male harbor seal (Phoca vitulina) produces broadband nonharmonic vocalizations underwater during the breeding season. In total, 120 vocalizations from six colonies were analyzed to provide a description of the acoustic structure and for the presence of geographic variation. The complex harbor seal vocalizations may be described by how the frequency bandwidth varies over time. An algorithm that identifies the boundaries between noise and signal from digital spectrograms was developed in order to extract a frequency bandwidth contour. The contours were used as inputs for multivariate analysis. The vocalizations' sound types (e.g., pulsed sound, whistle, and broadband nonharmonic sound) were determined by comparing the vocalizations' spectrographic representations with sound waves produced by known sound sources. Comparison between colonies revealed differences in the frequency contours, as well as some geographical variation in use of sound types. The vocal differences may reflect a limited exchange of individuals between the six colonies due to long distances and strong site fidelity. Geographically different vocal repertoires have potential for identifying discrete breeding colonies of harbor seals, but more information is needed on the nature and extent of early movements of young, the degree of learning, and the stability of the vocal repertoire. A characteristic feature of many vocalizations in this study was the presence of tonal-like introductory phrases that fit into the categories pulsed sound and whistles. The functions of these phrases are unknown but may be important in distance perception and localization of the sound source. The potential behavioral consequences of the observed variability may be indicative of adaptations to different environmental properties influencing determination of distance and direction and plausible different male mating tactics.

Acoustics↗

Enhancing Chinese tone recognition by manipulating amplitude envelope: implications for cochlear implants.

Tone recognition is important for speech understanding in tonal languages such as Mandarin Chinese. Cochlear implant patients are able to perceive some tonal information by using temporal cues such as periodicity-related amplitude fluctuations and similarities between the fundamental frequency (F0) contour and the amplitude envelope. The present study investigates whether modifying the amplitude envelope to better resemble the F0 contour can further improve tone recognition in multichannel cochlear implants. Chinese tone and vowel recognition were measured for six native Chinese normal-hearing subjects listening to a simulation of a four-channel cochlear implant speech processor with and without amplitude envelope enhancement. Two algorithms were proposed to modify the amplitude envelope to more closely resemble the F0 contour. In the first algorithm, the amplitude envelope as well as the modulation depth of periodicity fluctuations was adjusted for each spectral channel. In the second algorithm, the overall amplitude envelope was adjusted before multichannel speech processing, thus reducing any local distortions to the speech spectral envelope. The results showed that both algorithms significantly improved Chinese tone recognition. By adjusting the overall amplitude envelope to match the F0 contour before multichannel processing, vowel recognition was better preserved and less speech-processing computation was required. The results suggest that modifying the amplitude envelope to more closely resemble the F0 contour may be a useful approach toward improving Chinese-speaking cochlear implant patients' tone recognition.

Adult↗

The control of aerodynamics, acoustics, and perceptual characteristics during speech production.

One of the most important areas of study in speech motor control is the identification of control variables, the variables controlled by the nervous system during motor tasks. The current study examined two hypotheses regarding control variables in speech production: (1) pressure and resistance in the vocal tract are controlled, and (2) perceptual and acoustic accuracy are controlled. Aerodynamic and acoustic data were collected on 20 subjects in three conditions, normally (NT), with an open air pressure bleed tube in place (TWB), and with a closed bleed tube in place (TNB). The voice recordings collected from the speakers in the production study were used in the perceptual study. Results showed that oral pressure (Po) was significantly lower in the TWB condition than in the NT and TNB conditions. The Po in the TWB condition seemed to be related to maintenance of subglottal pressure (Ps). Examination of the perceptual and acoustic data indicated that perceptual accuracy for [a] was achieved by maintaining Ps to preserve a steady sound pressure level, fundamental frequency, and voicing. Overall, it appeared speakers controlled pressure in compensating, but for the ultimate goal of maintaining acoustic and perceptual accuracy.

Adult↗

Geographic variations in the whistles of spinner dolphins (Stenella longirostris) of the Main Hawai'ian Islands.

Geographic variations in the whistles of Hawai'ian spinner dolphins are discussed by comparing 27 spinner dolphin pods recorded in waters off the Islands of Kaua'i, O'ahu, Lana'i, and Hawai'i. Three different behavioral states, the number of dolphins observed in each pod, and ten parameters extracted from each whistle contour were considered by using clustering and discriminant function analyses. The results suggest that spinner dolphin pods in the Main Hawai'ian Islands share characteristics in approximately 48% of their whistles. Spinner dolphin pods had similar whistle parameters regardless of the island, location, and date when they were sampled and the dolphins' behavioral state and pod size. The term "whistle-specific subgroup" (WSS) was used to designate whistle groups with similar whistles parameters (which could have been produced in part by the same dolphins). The emission rate of whistles was higher when spinner dolphins were socializing than when they were traveling or resting, suggesting that whistles are mainly used during close-range interactions. Spinner dolphins also seem to vary whistle duration according to their general behavioral state. Whistle duration and the number of turns and steps of a whistle may be more important in delivering information at the individual level than whistle frequency parameters.

Animals↗

Robust and accurate fundamental frequency estimation based on dominant harmonic components.

This paper presents a new method for robust and accurate fundamental frequency (F0) estimation in the presence of background noise and spectral distortion. Degree of dominance and dominance spectrum are defined based on instantaneous frequencies. The degree of dominance allows one to evaluate the magnitude of individual harmonic components of the speech signals relative to background noise while reducing the influence of spectral distortion. The fundamental frequency is more accurately estimated from reliable harmonic components which are easy to select given the dominance spectra. Experiments are performed using white and babble background noise with and without spectral distortion as produced by a SRAEN filter. The results show that the present method is better than previously reported methods in terms of both gross and fine F0 errors.

Adult↗

Clear speech perception in acoustic and electric hearing.

When instructed to speak clearly for people with hearing loss, a talker can effectively enhance the intelligibility of his/her speech by producing "clear" speech. We analyzed global acoustic properties of clear and conversational speech from two talkers and measured their speech intelligibility over a wide range of signal-to-noise ratios in acoustic and electric hearing. Consistent with previous studies, we found that clear speech had a slower overall rate, higher temporal amplitude modulations, and also produced higher intelligibility than conversational speech. To delineate the role of temporal amplitude modulations in clear speech, we extracted the temporal envelope from a number of frequency bands and replaced speech fine-structure with noise fine-structure to simulate cochlear implants. Although both simulated and actual cochlear-implant listeners required higher signal-to-noise ratios to achieve normal performance, a 3-4 dB difference in speech reception threshold was preserved between clear and conversational speech for all experimental conditions. These results suggest that while temporal fine structure is important for speech recognition in noise in general, the temporal envelope carries acoustic cues that contribute to the clear speech intelligibility advantage.

Adult↗

Vocal tract resonances in singing: the soprano voice.

The vocal tract resonances of trained soprano singers were measured while they sang a range of vowels softly at different pitches. The measurements were made by broad band acoustic excitation at the mouth, which allowed the resonances of the tract to be measured simultaneously with and independently from the harmonics of the voice. At low pitch, when the lowest resonance frequency R1 exceeded f0, the values of the first two resonances R1 and R2 varied little with frequency and had values consistent with normal speech. At higher pitches, however, when fo exceeded the value of R1 observed at low pitch, R1 increased with f0 so that R1 was approximately equal to f0. R2 also increased over this high pitch range, probably as an incidental consequence of the tuning of R1. R3 increased slightly but systematically, across the whole pitch range measured. There was no evidence that any resonances are tuned close to harmonics of the pitch frequency except for R1 at high pitch. The variations in R1 and R2 at high pitch mean that vowels move, converge, and overlap their positions on the vocal plane (R2,R1) to an extent that implies loss of intelligibility.

Acoustics↗

Source localization in complex listening situations: selection of binaural cues based on interaural coherence.

In everyday complex listening situations, sound emanating from several different sources arrives at the ears of a listener both directly from the sources and as reflections from arbitrary directions. For localization of the active sources, the auditory system needs to determine the direction of each source, while ignoring the reflections and superposition effects of concurrently arriving sound. A modeling mechanism with these desired properties is proposed. Interaural time difference (ITD) and interaural level difference (ILD) cues are only considered at time instants when only the direct sound of a single source has non-negligible energy in the critical band and, thus, when the evoked ITD and ILD represent the direction of that source. It is shown how to identify such time instants as a function of the interaural coherence (IC). The source directions suggested by the selected ITD and ILD cues are shown to imply the results of a number of published psychophysical studies related to source localization in the presence of distracters, as well as in precedence effect conditions.

Acoustic Stimulation↗

The apparent immunity of high-frequency "transposed" stimuli to low-frequency binaural interference.

Discrimination of interaural temporal disparities (ITDs) was measured with either conventional or transposed "targets" centered at 4 kHz. The targets were presented either in the presence or absence of a simultaneously gated diotic noise centered at 500 Hz, the interferer. As expected, the presence of the low-frequency interferer resulted in substantially elevated threshold-ITDs for the conventional high-frequency stimuli. In contrast, these interference effects were absent for ITDs conveyed by the high-frequency transposed targets. The binaural interference effects observed with the conventional high-frequency stimuli were well accounted for, quantitatively, by the model described by Heller and Trahiotis [L. M. Heller and C. Trahiotis, J. Acoust. Soc. Am. 99, 3632-3637 (1996)]. The lack of binaural interference effects observed with the high-frequency transposed stimuli was not predicted by that model. It is suggested that transposed stimuli may be one of a class of stimuli that do not foster an obligatory combination of binaural information between low- and high-frequency regions. Under those conditions that do foster such an obligatory combination, one could still consider models of binaural interference, such as the one described in Heller and Trahiotis, to be valid descriptors of binaural processing.

Acoustic Stimulation↗