Search PubMed⌕ Search

Biomedical subjects

R Harald Baayen

Publications and source records attributed to R Harald Baayen.

14 recordsLinked to original sources

Lexical frequency and voice assimilation.

Acoustic duration and degree of vowel reduction are known to correlate with a word's frequency of occurrence. The present study broadens the research on the role of frequency in speech production to voice assimilation. The test case was regressive voice assimilation in Dutch. Clusters from a corpus of read speech were more often perceived as unassimilated in lower-frequency words and as either completely voiced (regressive assimilation) or, unexpectedly, as completely voiceless (progressive assimilation) in higher-frequency words. Frequency did not predict the voice classifications over and above important acoustic cues to voicing, suggesting that the frequency effects on the classifications were carried exclusively by the acoustic signal. The duration of the cluster and the period of glottal vibration during the cluster decreased while the duration of the release noises increased with frequency. This indicates that speakers reduce articulatory effort for higher-frequency words, with some acoustic cues signaling more voicing and others less voicing. A higher frequency leads not only to acoustic reduction but also to more assimilation.

Female↗

The nature of anterior negativities caused by misapplications of morphological rules.

This study investigates functional interpretations of left anterior negativities (LANs), a language-related electroencephalogram effect that has been found for syntactic and morphological violations. We focus on three possible interpretations of LANs caused by the replacement of irregular affixes with regular affixes: misapplication of morphological rules, mismatch of the presented form with analogy-based expectations, and mismatch of the presented form with stored representations. Event-related brain potentials were recorded during the visual presentation of existing and novel Dutch compounds. Existing compounds contained correct or replaced interfixes (dame + s + salons > damessalons vs. *dame + n + salons > *damensalons "women's hairdresser salons"), whereas novel Dutch compounds contained interfixes that were either supported or not supported by analogy to similar existing compounds (kruidenkelken vs. ?kruidskelken "herb chalices"); earlier studies had shown that interfixes are selected by analogy instead of rules. All compounds were presented with correct or incorrect regular plural suffixes (damessalons vs. *damessalonnen). Replacing suffixes or interfixes in existing compounds both led to increased (L)ANs between 400 and 700 msec without any evidence for different scalp distributions for interfixes and suffixes. There was no evidence for a negativity when manipulating the analogical support for interfixes in novel compounds. Together with earlier studies, these results suggest that LANs had been caused by the mismatch of the presented forms with stored forms. We discuss these findings with respect to the single/dual-route debate of morphology and LANs found for the misapplication of syntactic rules.

Adolescent↗

Articulatory planning is continuous and sensitive to informational redundancy.

This study investigates the relationship between word repetition, predictability from neighbouring words, and articulatory reduction in Dutch. For the seven most frequent words ending in the adjectival suffix -lijk, 40 occurrences were randomly selected from a large database of face-to-face conversations. Analysis of the selected tokens showed that the degree of articulatory reduction (as measured by duration and number of realized segments) was affected by repetition, predictability from the previous word and predictability from the following word. Interestingly, not all of these effects were significant across morphemes and tar-get words. Repetition effects were limited to suffixes, while effects of predictability from the previous word were restricted to the stems of two of the seven target words. Predictability from the following word affected the stems of all target words equally, but not all suffixes. The implications of these findings for models of speech production are discussed.

Humans↗

Frequency effects in compound production.

Four experiments investigated the role of frequency information in compound production by independently varying the frequencies of the first and second constituent as well as the frequency of the compound itself. Pairs of Dutch noun-noun compounds were selected such that there was a maximal contrast for one frequency while matching the other two frequencies. In a position-response association task, participants first learned to associate a compound with a visually marked position on a computer screen. In the test phase, participants had to produce the associated compound in response to the appearance of the position mark, and we measured speech onset latencies. The compound production latencies varied significantly according to factorial contrasts in the frequencies of both constituting morphemes but not according to a factorial contrast in compound frequency, providing further evidence for decompositional models of speech production. In a stepwise regression analysis of the joint data of Experiments 1-4, however, compound frequency was a significant nonlinear predictor, with facilitation in the low-frequency range and a trend toward inhibition in the high-frequency range. Furthermore, a combination of structural measures of constituent frequencies and entropies explained significantly more variance than a strict decompositional model, including cumulative root frequency as the only measure of constituent frequency, suggesting a role for paradigmatic relations in the mental lexicon.

Humans↗

Shifting paradigms: gradient structure in morphology.

Morphology is the study of the internal structure of words. A vigorous ongoing debate surrounds the question of how such internal structure is best accounted for: by means of lexical entries and deterministic symbolic rules, or by means of probabilistic subsymbolic networks implicitly encoding structural similarities in connection weights. In this review, we separate the question of subsymbolic versus symbolic implementation from the question of deterministic versus probabilistic structure. We outline a growing body of evidence, mostly external to the above debate, indicating that morphological structure is indeed intrinsically graded. By allowing probability into the grammar, progress can be made towards solving some long-standing puzzles in morphological theory.

Humans↗

Lexical frequency and acoustic reduction in spoken Dutch.

This study investigates the effects of lexical frequency on the durational reduction of morphologically complex words in spoken Dutch. The hypothesis that high-frequency words are more reduced than low-frequency words was tested by comparing the durations of affixes occurring in different carrier words. Four Dutch affixes were investigated, each occurring in a large number of words with different frequencies. The materials came from a large database of face-to-face conversations. For each word containing a target affix, one token was randomly selected for acoustic analysis. Measurements were made of the duration of the affix as a whole and the durations of the individual segments in the affix. For three of the four affixes, a higher frequency of the carrier word led to shorter realizations of the affix as a whole, individual segments in the affix, or both. Other relevant factors were the sex and age of the speaker, segmental context, and speech rate. To accommodate for these findings, models of speech production should allow word frequency to affect the acoustic realizations of lower-level units, such as individual speech sounds occurring in affixes.

Humans↗

Prosodic cues for morphological complexity: the case of Dutch plural nouns.

It has recently been shown that listeners use systematic differences in vowel length and intonation to resolve ambiguities between onset-matched simple words (Davis, Marslen-Wilson, & Gaskell, 2002; Salverda, Dahan, & McQueen, 2003). The present study shows that listeners also use prosodic information in the speech signal to optimize morphological processing. The precise acoustic realization of the stem provides crucial information to the listener about the morphological context in which the stem appears and attenuates the competition between stored inflectional variants. We argue that listeners are able to make use of prosodic information, even though the speech signal is highly variable within and between speakers, by virtue of the relative invariance of the duration of the onset. This provides listeners with a baseline against which the durational cues in a vowel and a coda can be evaluated. Furthermore, our experiments provide evidence for item-specific prosodic effects.

Cues↗

Putting the bits together: an information theoretical perspective on morphological processing.

In this study we introduce an information-theoretical formulation of the emergence of type- and token-based effects in morphological processing. We describe a probabilistic measure of the informational complexity of a word, its information residual, which encompasses the combined influences of the amount of information contained by the target word and the amount of information carried by its nested morphological paradigms. By means of re-analyses of previously published data on Dutch words we show that the information residual outperforms the combination of traditional token- and type-based counts in predicting response latencies in visual lexical decision, and at the same time provides a parsimonious account of inflectional, derivational, and compounding processes.

Decision Making↗

Morphological family size in a morphologically rich language: the case of Finnish compared with Dutch and Hebrew.

Finnish has a very productive morphology in which a stem can give rise to several thousand words. This study presents a visual lexical decision experiment addressing the processing consequences of the huge productivity of Finnish morphology. The authors observed that in Finnish words with larger morphological families elicited shorter response latencies. However, in contrast to Dutch and Hebrew, it is not the complete morphological family of a complex Finnish word that codetermines response latencies but only the subset of words directly derived from the complex word itself. Comparisons with parallel experiments using translation equivalents in Dutch and Hebrew showed substantial cross-language predictivity of family size between Finnish and Dutch but not between Finnish and Hebrew, reflecting the different ways in which the Hebrew and Finnish morphological systems contribute to the semantic organization of concepts in the mental lexicon.

Cross-Cultural Comparison↗

Affixal homonymy triggers full-form storage, even with inflected words, even in a morphologically rich language.

This paper investigates whether affixal homonymy, the phenomenon that one affix form serves two or more semantic/syntactic functions, affects lexical processing of inflected words in a similar way for a morphologically rich language such as Finnish as for morphologically restricted languages such as Dutch and English. For the latter two languages, there is evidence that affixal homonymy triggers full-form storage for inflected words (Bertram, R., Schreuder, R., and Baayen, R. H. (in press). The balance of storage and computation in morphological processing: the role of word formation type, affixal homonymy, and productivity. Journal of Experimental Psychology: Learning, Memory, and Cognition; Sereno and Jongman (1997). Processing of English inflectional morphology. Memory and Cognition, 25, 425-437). Two visual lexical decision experiments show the same pattern for Finnish. Apparently, the substantially richer morphology in Finnish does not prevent full-form storage for inflected words when the affix is homonymic.

Cognition↗

The subjects as a simple random effect fallacy: subject variability and morphological family effects in the mental lexicon.

This is a methodological study addressing the appropriateness of standard by-subject and by-item averaging procedures for the analysis of repeated-measures designs. By means of a reanalysis of published data (Schreuder & Baayen, 1997), using random regression models, we present a proof of existence of systematic variability between participants that is ignored in the standard psycholinguistic analytical procedures. By applying linear mixed effects modeling (Pinheiro & Bates, 2000), we call attention to the potential lack of power of the by-subject and by-item analyses, which in this case study fail to reveal the coexistence of a facilitatory family size effect and an inhibitory family frequency effect in visual and auditory lexical processing.

Auditory Perception↗

The processing and representation of Dutch and English compounds: peripheral morphological and central orthographic effects.

In this study, we use the association between various measures of the morphological family and decision latencies to reveal the way in which the components of Dutch and English compounds are processed. The results show that for constituents of concatenated compounds in both languages, a position-related token count of the morphological family plays a role, whereas English open compounds show an effect of a type count, similar to the effect of family size for simplex words. When Dutch compounds are written with an artificial space, they reveal no effect of type count, which shows that the differential effect for the English open compounds is not superficial. The final experiment provides converging evidence for the lexical consequences of the space in English compounds. Decision latencies for English simplex words are better predicted from counts of the morphological family that include concatenated and hyphenated but not open family members.

Cognition↗

Linking elements in Dutch noun-noun compounds: constituent families as analogical predictors for response latencies.

This study addresses the choice of linking elements in novel Dutch noun-noun compounds. Previous off-line experiments (Krott, Baayen, & Schreuder, 2001) revealed that this choice can be predicted analogically on the basis of the distribution of linking elements in the left and right constituent families, i.e., the set of existing compounds that share the left (or right) constituent with the target compound. The present study replicates the observed graded analogical effects under time pressure, using an on-line decision task. Furthermore, the analogical support of the left constituent family predicts response latencies. We present an implemented interactive activation network model that accounts for the experimental data.

Cognition↗

Do type and token effects reflect different mechanisms? Connectionist modeling of Dutch past-tense formation and final devoicing.

In this paper, we show that both token and type-based effects in lexical processing can result from a single, token-based, system, and therefore, do not necessarily reflect different levels of processing. We report three Simple Recurrent Networks modeling Dutch past-tense formation. These networks show token-based frequency effects and type-based analogical effects closely matching the behavior of human participants when producing past-tense forms for both existing verbs and pseudo-verbs. The third network covers the full vocabulary of Dutch, without imposing predefined linguistic structure on the input or output words.

Computer Simulation↗