1358 norm sets
The present article provides normative measures for 400 pictured objects (Cycowicz et al., 1997) viewed by Portuguese speaking Brazilian University students and 5-7 year-old children. Name agreement, familiarity and visual complexity ratings were obtained. These variables have been shown to be important for the selection of adequate stimuli for cognitive studies. Children's name agreement was lower than that of adults. The children also failed to provide adequate modal names for 103 concepts, rated drawings as less familiar and less complex, and chose shorter names for pictures. The differences in ratings between adults and children were higher than those observed in the literature employing smaller picture sets. The pattern of correlations among measures observed in the present study was consistent with previous reports, supporting the usefulness of the 400 picture set as a tool for cognitive research in different cultures and ages.
With a view to designing a speaker-independent large vocabulary recognition system, we evaluate a vector quantization approach for speaker adaptation. Only one speaker (the reference speaker) pronounces the application vocabulary. He also pronounces a small vocabulary called the adaptation vocabulary. Each new speaker then merely pronounces the adaptation vocabulary. We have compared two adaptation methods, establishing a correspondence between the codebooks of the reference and the new speakers, on a 20-speaker database with a 104-word application vocabulary. Method I uses a transposed codebook to represent the new speaker during the recognition process, whereas Method II uses a codebook which is obtained by clustering analysis on the NS's pronunciation of the adaptation vocabulary. The adaptation vocabulary contains 136 words. Comparison of the performance of the two methods shows that a new speaker's codebook is not necessary to represent the new speaker. Consequently we have used the first method to perform tests with a 5000-word application vocabulary, and a 4-speaker database. The adaptation is still efficient (the mean improvement is about 14{\%}), even if the relative improvement is 30{\%} compared to 56{\%} obtained in the 104-word application experiment. Further experiments show that the recognition accuracy can be improved by increasing the adaptation vocabulary size and the codebook size. {\textcopyright} 1991.
Newly measured rating norms provide a database of emotion-related dimensions for 524 French trait words. Measures include valence, approach/avoidance tendencies associated with the trait, possessor- and other-relevance of the trait, and discrete emotions conveyed by the trait (i.e., anger, disgust, fear, happiness, and sadness). The normative data were obtained from 328 participants and were revealed to be stable across samples and gender. These data go beyond a dimensional structure and consider more fine-grained descriptions such as the categorical emotions, as well as the perspective of the evaluator conveyed by the traits. They should thus be particularly useful for researchers interested in emotion or in the emotional dimension of cognition, action, or personality. The database is available as supplementary material.
Indicators of letter visual similarity have been used for controlling the design of empirical and neuropsychological studies and for rigorously determining the factors that underlie reading ability and literacy acquisition. Additionally, these letter similarity/confusability matrices have been useful for studies examining more general aspects of human cognition, such as perception. Despite many letter visual-similarity matrices being available, they all have two serious limitations if they are to be used by researchers in the reading domain: (1) They have been constructed using atypical reading data obtained from speeded reading-aloud tasks and/or under degraded presentation conditions; (2) they only include letters from the English alphabet. Although some letter visual-similarity matrices have been constructed using data gathered from normal reading conditions, these either are based on old fonts, which may not resemble the letters found in modern print, or were never published. For the first time, this article presents a comprehensive letter visual-similarity/confusability matrix that has been constructed based on untimed responses to clearly presented upper- and lowercase letters that are present in many languages that use Latin-based alphabets, including Catalan, Dutch, English, French, Galician, German, Italian, Portuguese, and Spanish. Such a matrix will be useful for researchers interested in the processes underpinning reading and literacy acquisition.
Conducted 9 experiments with a total of 663 undergraduates using the technique of priming to study the nature of the cognitive representation generated by superordinate semantic category names. In Exp I, norms for the internal structure of 10 categories were collected. In Exps II, III, and IV, internal structure was found to affect the perceptual encoding of physically identical pairs of stimuli, facilitating responses to physically identical good members and hindering responses to identical poor members of a category. Exps V and VI showed that the category name did not generate a physical code (e.g., lines or angles), but rather affected perception of the stimuli at the level of meaning. Exps VII and VIII showed that while the representation of the category name which affected perception contained a depth meaning common to words and pictures which enabled Ss to prepare for either stimulus form within 700 msec, selective reduction of the interval between prime and stimulus below 700 msec revealed differentiation of the coding of meaning in preparation for actual perception. Exp IX suggested that good examples of semantic categories are not physiologically determined, as the effects of the internal structure of semantic categories on priming (unlike the effects for color categories) could be eliminated by long practice.
Ratings of pleasantness (PL) on a 7-point scale and of association value (a′) on a 5-point scale are reported for 101 two-syllable nouns. The ratings were obtained from two samples of 100 women each and two samples of 100 men each. Sizable differences were obtained between words on both scales. For women and men respectively, PL and a′ were correlated .570 and .585; PL and printed frequency were correlated .233 and .207; frequency and a′ were correlated .533 and .764. Women's and men's ratings correlated .973 for PL and .899 for a′. {\textcopyright} 1969 Academic Press Inc. All rights reserved.
The aim of the present study was to provide normative data for the Croatian language using 346 visually presented objects (Cycowicz, Friedman, Rothstein, {\&} Snodgrass Journal of Experimental Child Psychology 65:171-237, 1997; Roach, Schwartz, Martin, Grewal, {\&} Brecher Clinical Aphasiology 24:121-133, 1996; Snodgrass {\&} Vanderwart Journal of Experimental Psychology: Human Learning and Memory 6:174-215, 1980). Picture naming was standardized according to seven variables: naming latency, name agreement, familiarity, visual complexity, word length, number of syllables, and word frequency. The descriptive statistics and correlation pattern of the variables collected in the present study were consistent with normative studies in other languages. These normative data for pictorial stimuli named by young healthy Croatian native speakers will be useful in studies of perception, language, and memory, as well as for preoperative and intraoperative mapping of speech and language brain areas.
The present study provides a French child database containing a large corpus of words (N = 600) that were rated on emotional valence (positive, neutral, and negative) by French children differing in both age (5, 7, and 9 years old) and sex (girls and boys). Good response reliability was observed in each of the three age groups. The results showed some age differences in the children's ratings. With increasing age, the percentage of words rated positive decreased, whereas the percentage of neutral words increased and the percentage of negative words remained stable. Our study did not reveal marked differences across sex groups. The database compiled here should become a useful tool for experimental studies in which verbal material is used with children. It would be worthwhile in future research to study how children process emotional words and also to control the emotional variable in the same way as other linguistic variables in the experimental design. The norms from this study may be downloaded from brm.psychonomic-journals.org/content/supplemental.
Body-object interaction (BOI) assesses the ease with which a human body can physically interact with a word's referent. Recent research has shown that BOI influences visual word recognition processes in such a way that responses to high-BOI words (e.g., couch) are faster and less error prone than responses to low-BOI words (e.g., cliff). Importantly, the high-BOI words and the low-BOI words that were used in those studies were matched on imageability. In the present study, we collected BOI ratings for a large set of words. BOI ratings, on a 1-7 scale, were obtained for 1,618 monosyllabic nouns. These ratings allowed us to test the generalizability of BOI effects to a large set of items, and they should be useful to researchers who are interested in manipulating or controlling for the effects of BOI. The body-object interaction ratings for this study may be downloaded from the Psychonomic Society's Archive of Norms, Stimuli, and Data, www.psychonomic.org/archive.
An online calculator was developed (www.bncdnet.ku.edu/cml/info{\_}ccc.vi) to compute phonotactic probability--the likelihood of occurrence of a sound sequence--and neighborhood density--the number of phonologically similar words--on the basis of child corpora of American English (Kolson, 1960; Moe, Hopkins, {\&} Rush, 1982) and to compare its results to those of an adult calculator. Phonotactic probability and neighborhood density were computed for a set of 380 nouns (Fenson et al., 1993) using both the child and adult corpora. The child and adult raw values were significantly correlated. However, significant differences were detected. Specifically, child phonotactic probability was higher than adult phonotactic probability, especially for high-probability words, and child neighborhood density was lower than adult neighborhood density, especially for words with high-density neighborhoods. These differences were reduced or eliminated when relative measures (i.e., z scores) were used. Suggestions are offered regarding which values to use in future research.
Although there are many well-characterized affective visual stimuli sets available to researchers, there are few auditory sets available. Those auditory sets that are available have been characterized primarily according to one of two major theories of affect: dimensional or categorical. Current trends have attempted to utilize both theories to more fully understand emotional processing. As such, stimuli that have been thoroughly characterized according to both of these approaches are exceptionally useful. In an effort to provide researchers with such a stimuli set, we collected descriptive data on the International Affective Digitized Sounds (IADS), identifying which discrete categorical emotions are elicited by each sound. The IADS is a database of 111 sounds characterized along the affective dimensions of valence, arousal, and dominance. Our data complement these characterizations of the IADS, allowing researchers to control for or manipulate stimulus properties in accordance with both theories of affect, providing an avenue for further integration of these perspectives. Related materials may be downloaded from the Psychonomic Society Web archive at www.psychonomic.org/archive.
Word associations to each of the 26 letters of the alphabet were obtained under procedures of single association or continued associations for both upper- and lower-case letters. The results showed a significant relationship between m values and measures of frequency of letters, preferences for letters, and vocal reaction time to letters. The data also showed that m values for each letter were stable within a session. Analyses of the most frequent associations showed a high degree of consistency among the common associations for the single and continued instructional procedures and upper- and lower-case stimulus presentations of the letters. {\textcopyright} 1965 Academic Press Inc. All rights reserved.
The Affective Norms for English Words (ANEW) are a commonly used set of 1,034 words characterized on the affective dimensions of valence, arousal, and dominance. Traditionally, studies of affect have used stimuli characterized along either affective dimensions or discrete emotional categories, but much current research draws on both of these perspectives. As such, stimuli that have been thoroughly characterized according to both of these approaches are exceptionally useful. In an effort to provide researchers with such a characterization of stimuli, we have collected descriptive data on the ANEW to identify which discrete emotions are elicited by each word in the set. Our data, coupled with previous characterizations of the dimensional aspects of these words, will allow researchers to control for or manipulate stimulus properties in accordance with both dimensional and discrete emotional views, and provide an avenue for further integration of these two perspectives. Our data have been archived at www.psychonomic.org/archive/.
- SP{\'{I}}{\v{S}} FONOLOGIE, MOORY ATD$\backslash$r$\backslash$nOn the basis of the lexical corpus created by Amano and Kondo (2000), using the Asahi newspaper, the present study provides frequencies of occurrence for units of Japanese phonemes, morae, and syllables. Among the five vowels, /a/ (23.42{\%}), /i/ (21.54{\%}), /u/ (23.47{\%}), and /o/ (20.63{\%}) showed similar frequency rates, whereas /e/ (10.94{\%}) was less frequent. Among the 12 consonants, /k/ (17.24{\%}), /t/ (15.53{\%}), and /r/ (13.11{\%}) were used often, whereas /p/ (0.60{\%}) and /b/ (2.43{\%}) appeared far less frequently. Among the contracted sounds, /sj/ (36.44{\%}) showed the highest frequency, whereas /mj/ (0.27{\%}) rarely appeared. Among the five long vowels, /aR/ (34.4{\%}) was used most frequently, whereas /uR/ (12.11{\%}) was not used so often. The special sound /N/ appeared very frequently in Japanese. The syllable combination /k/+V+/N/ (19.91{\%}) appeared most frequently among syllabic combinations with the nasal /N/. The geminate (or voiceless obstruent) /Q/, when placed before the four consonants /p/, /t/, /k/, and /s/, appeared 98.87{\%} of the time, but the remaining 1.13{\%} did not follow the definition. The special sounds /R/, /N/, and /Q/ seem to appear very frequently in Japanese, suggesting that they are not special in terms of frequency counts. The present study further calculated frequencies for the 33 newly and officially listed morae/syllables, which are used particularly for describing alphabetic loanwords. In addition, the top 20 bi-mora frequency combinations are reported. Files of frequency indexes may be downloaded from the Psychonomic Society Web archive at http://www.psychonomic.org/archive/.
The present study reports descriptive normative measures for 245 Italian verbal idiomatic expressions. For each of the idiomatic expressions the following variables are reported: Length, Knowledge, Familiarity, Age of Acquisition, Predictability, Syntactic flexibility, Literality and Compositionality. Syntactic flexibility was assessed using five syntactic operations: adverb insertion, adjective insertion, left dislocation, passive and movement. The psycholinguistic relevance of each dimension, their measures and the correlations among them are provided and discussed. The databases are freely available for down-loading from the Psychonomic Society Web archive at www.psychonomic.org/archive/.
In 1981, the Japanese government published a list of the 1,945 basic Japanese kanji (Jooyoo Kanji-hyo), including specifications of pronunciation. This list was established as the standard for kanji usage in print. The database for 1,945 basic Japanese kanji provides 30 cells that explain in detail the various characteristics of kanji. Means, standard deviations, distributions, and information related to previous research concerning these kanji are provided in this paper. The database is saved as a Microsoft Excel 2000 file for Windows. This kanji database is accessible on the Web site of the Oxford Text Archive, Oxford University (http://ota.ahds.ac.uk). Using this database, researchers and educators will be able to conduct planned experiments and organize classroom instruction on the basis of the known characteristics of selected kanji.
The lexical database dlexDB supplies in form of an online database frequency-based norms of numerous process-related word properties for psychological and linguistic research. These values include well known variables such as printed frequency of word form and lemma as documented also in CELEX (Baayen, Piepenbrock und Gulikers, 1995). In addition, we compute new values like frequencies based on syllables, and morphemes as well as frequencies of character chains, and multiple word combinations. The statistics are based on the Kernkorpus des Digitalen Wrterbuchs der deutschen Sprache (DWDS) with over 100 million running words. We illustrate the validity of these norms with new results about fixation durations in sentence reading.
Many cognitive psychological, computational, and neuropsychological approaches to the organisation of semantic memory have incorporated the idea that concepts are, at least partly, represented in terms of their fine-grained features. We asked 20 normal volunteers to provide properties of 64 concrete items, drawn from living and nonliving categories, by completing simple sentence stems (e.g., an owl is {\_}{\_}, has {\_}{\_}, can{\_}{\_}). At a later date, the same participants rated the same concepts for prototypicality and familiarity. The features generated were classified as to type of knowledge (sensory, functional, or encyclopaedic), and also quantified with regard to both dominance (the number of participants specifying that property for that concept) and distinctiveness (the proportion of exemplars within a conceptual category of which that feature was considered characteristic). The results demonstrate that rated prototypicality is related to both the familiarity of the concept and its distance from the average of the exemplars within the same category (the category centroid). The feature database was also used to replicate, resolve, and extend a variety of previous observations on the structure of semantic representations. Specifically, the results of our analyses (1) resolve two conflicting claims regarding the relative ratio of sensory to other kinds of attributes in living vs. nonliving concepts; (2) offer new information regarding the types of features-across different domains-that distinguish concepts from their category coordinates; and (3) corroborate some previous claims of higher intercorrelations between features of living things than those of artefacts.
Semantic differential (SD) factor scores on the Evaluation, Activity, and Potency dimensions are presented for 1,000 most frequently used English words. Also given are the standard errors of the factor scores, the results of several reliability studies, and a listing (for all words) of 3 types of derived scores: polarizations, n Affiliation contents, n Achievement contents. Test-ing procedures and statistics on the sample of raters are detailed. Some uses of the dictionary are suggested, and an example of its use in a study of motivation is presented including empirical results. Conditions favoring further cumulation of SD data are discussed. THE semantic differential (SD) has proven to be an accurate instrument for recording affective associations of stim-uli, particularly to the extent that such as-sociations are culturally or subculturally denned so that measurements may be aver-aged over groups of individuals (Norman, 1959). In a wide variety of studies, includ-ing many involving cross-cultural samples of raters, it has been demonstrated that affective judgments on bipolar adjective scales reliably resolve into three major dimensions or factors which Osgood has named Evaluation, Activity, and Potency 'This paper is part of a doctoral dissertation submitted to the
As researchers explore the complexity of memory and language hierarchies, the need to expand normed stimulus databases is growing. Therefore, we present 1,808 words, paired with their features and concept-concept information, that were collected using previously established norming methods (McRae, Cree, Seidenberg, {\&} McNorgan Behavior Research Methods 37:547-559, 2005). This database supplements existing stimuli and complements the Semantic Priming Project (Hutchison, Balota, Cortese, Neely, Niemeyer, Bengson, {\&} Cohen-Shikora 2010). The data set includes many types of words (including nouns, verbs, adjectives, etc.), expanding on previous collections of nouns and verbs (Vinson {\&} Vigliocco Journal of Neurolinguistics 15:317-351, 2008). We describe the relation between our and other semantic norms, as well as giving a short review of word-pair norms. The stimuli are provided in conjunction with a searchable Web portal that allows researchers to create a set of experimental stimuli without prior programming knowledge. When researchers use this new database in tandem with previous norming efforts, precise stimuli sets can be created for future research endeavors.
We have developed a set of naming and recognition tests for evaluating the retrieval of lexical and conceptual knowledge for actions. As a first step, normative information about 280 items was collected for the following variables: (1) the naming responses elicited by each item, (2) the degree to which the image of each item agreed with a target name, (3) the familiarity to each depicted action, and (4) the visual complexity of each item. This information was used to develop administration and scoring procedures for a standardized test of action naming. The effectiveness and reliability of these procedures were evaluated in a second experiment. In a third experiment, five tests were developed to probe the retrieval of conceptual knowledge: (1) independently of the production of a naming response, (2) in response to pictorial and nonpictorial stimuli, (3) in terms of the attributes associated with specific actions, and (4) in terms of similarities and differences between various actions.
Semantic ambiguity is typically measured by sum-ming the number of senses or dictionary definitions that a word has. Such measures are somewhat subjective and may not adequately capture the full extent of variation in word meaning, particularly for polysemous words that can be used in many different ways, with subtle shifts in meaning. Here, we describe an alternative, computationally derived measure of ambiguity based on the proposal that the meanings of words vary continuously as a function of their contexts. On this view, words that appear in a wide range of contexts on diverse topics are more variable in meaning than those that appear in a restricted set of similar contexts. To quantify this variation, we performed latent semantic analysis on a large text corpus to estimate the semantic similarities of different linguistic contexts. From these estimates, we calculated the degree to which the different contexts associated with a given word vary in their meanings. We term this quantity a word's semantic diversity (SemD). We suggest that this approach provides an objective way of quantifying the subtle, context-dependent variations in word meaning that are often present in language. We demonstrate that SemD is correlated with other measures of ambiguity and contextual variability, as well as with frequency and imageability. We also show that SemD is a strong predictor of performance in semantic judgments in healthy individuals and in patients with semantic deficits, accounting for unique variance beyond that of other predictors. SemD values for over 30,000 English words are provided as supplementary materials.
Speeded naming and lexical decision data for 1,661 target words following related and unrelated primes were collected from 768 subjects across four different universities. These behavioral measures have been integrated with demographic information for each subject and descriptive characteristics for every item. Subjects also completed portions of the Woodcock-Johnson reading battery, three attentional control tasks, and a circadian rhythm measure. These data are available at a user-friendly Internet-based repository ( http://spp.montana.edu ). This Web site includes a search engine designed to generate lists of prime-target pairs with specific characteristics (e.g., length, frequency, associative strength, latent semantic similarity, priming effect in standardized and raw reaction times). We illustrate the types of questions that can be addressed via the Semantic Priming Project. These data represent the largest behavioral database on semantic priming and are available to researchers to aid in selecting stimuli, testing theories, and reducing potential confounds in their studies.
Factors affecting word retrieval were compared in a timed picture-naming paradigm for 520 drawings of objects. In prior timed and untimed studies by Snodgrass
The combining of individual concepts to form an emergent concept is a fundamental aspect of language, yet much less is known about it than about processing isolated words or sentences. To facilitate research on conceptual combination, we provide meaningfulness ratings for a large set of (2,160) noun-noun pairs. Half of these pairs (1,080) are reversed versions of the other half (e.g., SKI JACKET and JACKET SKI), to facilitate the comparison of successful and unsuccessful conceptual combination independently of constituent lexical items. The computer code used for obtaining these ratings through a Web interface is provided. To further enhance the usefulness of this resource, ancillary measures obtained from other sources are also provided for each pair. These measures include associate production norms, contextual relatedness in terms of latent semantic analysis distance, total number of letters, phrase-level usage frequency, and word-level usage frequency summed across the words in each pair. Results of correlation and regression analyses are also provided for a quantitative description of the stimulus set. A subset of these stimuli was used to identify neural correlates of successful conceptual combination Graves, Binder, Desai, Conant, {\&} Seidenberg, (NeuroImage 53:638-646, 2010). The stimuli can be used in other research and also provide benchmark data for evaluating the effectiveness of computational algorithms for predicting meaningfulness of noun-noun pairs.
An operational definition of abstractness in nouns was constructed by using the human discriminative response to identify two points on a scale of abstractness. This scale, consisting of 490 'abstract' and 571 'concrete' nouns, was found to have adequate reliability. When the scale was manipulated as an independent variable, the effect of abstractness on short-term recognition memory was highly significant, 'abstract' nouns being less well remembered than 'concrete' nouns. Frequency was found to be pertinent variable, independent of abstractness, very frequent nouns being less well remembered than some-what rarer nouns.
We make available word-by-word self-paced reading times and eye-tracking data over a sample of English sentences from narrative sources. These data are intended to form a gold standard for the evaluation of computational psycholinguistic models of sentence comprehension in English. We describe stimuli selection and data collection and present descriptive statistics, as well as comparisons between the two sets of reading times.
In order to provide a reliable measure of the similarity of uppercase English letters, a confusion matrix based on 1,200 presentations of each letter was established. To facilitate an analysis of the perceived structural characteristics, the confusion matrix was decomposed according to Luce's choice model into a symmetrical similarity matrix and a response bias vector. The underlying structure of the similarity matrix was assessed with both a hierarchical clustering and a multidimensional scaling procedure. This data is offered to investigators of visual information processing as a valuable tool for controlling not only the overall similarity of the letters in a study, but also their similarity on individual feature dimensions.
A sample of 100 college students ranked the alphabet according to their preference for the appearance of the capital letter. Rankings are presented for the total sample, and for subgroups based on age and sex. Coefficients of concordance among judges are low, but the rankings for the total sample and the age and sex subsamples appear to be quite reliable.
Researchers concerned with the development of cognitive functions are in need of standardized material that can be used with both adults and children. The present article provides normative measures for 400 line drawings viewed by 5- and 6-year-old children. The three variables obtained - name agreement, familiarity, and visual complexity - are important because of their potential effect on memory and other cognitive processes. The normative data collected in the present study indicate that young children are different from adults in both the name most frequently assigned and the number of alternative names provided. The alternative names given by the children are either coordinate names or names of objects that are visually similar to the pictured object. In addition, the failure (to name) rate is higher among young children compared to adults. Thus, we conclude that unequivocal interpretation of age-related differences in cognitive functions can be made only when age-appropriate pictorial stimuli are chosen. {\textcopyright} 1997 Academic Press.
Data from parent reports on 1,803 children--derived from a normative study of the MacArthur Communicative Development Inventories (CDIs)--are used to describe the typical course and the extent of variability in major features of communicative development between 8 and 30 months of age. The two instruments, one designed for 8-16-month-old infants, the other for 16-30-month-old toddlers, are both reliable and valid, confirming the value of parent reports that are based on contemporary behavior and a recognition format. Growth trends are described for children scoring at the 10th-, 25th-, 50th-, 75th-, and 90th-percentile levels on receptive and expressive vocabulary, actions and gestures, and a number of aspects of morphology and syntax. Extensive variability exists in the rate of lexical, gestural, and grammatical development. The wide variability across children in the time of onset and course of acquisition of these skills challenges the meaningfulness of the concept of the modal child. At the same time, moderate to high intercorrelations are found among the different skills both concurrently and predictively (across a 6-month period). Sex differences consistently favor females; however, these are very small, typically accounting for 1{\%}-2{\%} of the variance. The effects of SES and birth order are even smaller within this age range. The inventories offer objective criteria for defining typicality and exceptionality, and their cost effectiveness facilitates the aggregation of large data sets needed to address many issues of contemporary theoretical interest. The present data also offer unusually detailed information on the course of development of individual lexical, gestural, and grammatical items and features. Adaptations of the CDIs to other languages have opened new possibilities for cross-linguistic explorations of sequence, rate, and variability of communicative development.
Normative data on the objective age of acquisition (AoA) for 286 Russian words are presented in this article. In addition, correlations between the objective AoA and subjective ratings, name agreement, picture name agreement, imageability, familiarity, word frequency, and word length are provided, as are correlations between the objective AoA and two measures of exemplar dominance (exemplar generation frequency and the number of times an exemplar was named first). The correlations between the aforementioned variables are generally consistent with the correlations reported in other normative studies. The objective AoA data are highly correlated with the subjective AoA ratings, whereas the correlations between the objective AoA and other psycholinguistic variables are moderate. The correlations between the objective AoA of Russian words and similar data for other languages are moderately high. The complete word norms may be downloaded from supplementary material.
Malay, a language spoken by 250 million people, has a shallow alphabetic orthography, simple syllable structures, and transparent affixation--characteristics that contrast sharply with those of English. In the present article, we first compare the letter-phoneme and letter-syllable ratios for a sample of alphabetic orthographies to highlight the importance of separating language-specific from language-universal reading processes. Then, in order to develop a better understanding of word recognition in orthographies with more consistent mappings to phonology than English, we compiled a database of lexical variables (letter length, syllable length, phoneme length, morpheme length, word frequency, orthographic and phonological neighborhood sizes, and orthographic and phonological Levenshtein distances) for 9,592 Malay words. Separate hierarchical regression analyses for Malay and English revealed how the consistency of orthography-phonology mappings selectively modulates the effects of different lexical variables on lexical decision and speeded pronunciation performance. The database of lexical and behavioral measures for Malay is available at http://brm.psychonomic-journals.org/content/supplemental.
The present study provides Canadian French normative data for 388 line drawings from the European Picture Pool for Oral Naming (Protocole europ{\'{e}}en de d{\'{e}}nomination orale d'images; PEDOI; Kremin et al., 2003). One hundred eighty subjects were equally distributed for age group (18-39,40-59, 60-85), educational level (low, high), and sex. They rated pictures of objects on age of acquisition, name agreement, familiarity, and visual complexity. Syllable length and word frequency were also taken into account. The present study suggests that age of acquisition and name agreement show significant age-related differences. These results show that unequivocal interpretation of age-related differences can be made when age-appropriate norms are used.
Stimulus material for studying object-directed actions is needed in different research contexts, such as action observation, action memory, and imitation. Action items have been generated many times in individual laboratories across the world, but they are used in very few experiments. For future studies in the field, it would be worthwhile to have a larger set of action stimulus material available to a broader research community. Some smaller action databases have already been published, but those often focus on psycholinguistic parameters and static action stimuli. With this article, we introduce an action database with dynamic action stimuli. The database contains action descriptions of 1,754 object-directed actions that have been rated for familiarity in Germany and in China. For 784 of these actions, action video clips are available. With the use of our database, it is possible to identify actions that differ in familiarity between Western and Eastern cultures. This variable may be of interest to some researchers in the field, since it has been shown that familiarity influences action information processing. Action descriptions are listed and categorized in tables that can be downloaded, along with the corresponding video clips, as supplemental material.
We collected number-of-translation norms on 562 Dutch-English translation pairs from several previous studies of cross-language processing. Participants were highly proficient Dutch-English bilinguals. Form and semantic similarity ratings were collected on the 1,003 possible translation pairs. Approximately 40{\%} of the translations were rated as being similar across languages with respect to spelling/sound (i.e., they were cognates). Approximately 45{\%} of the translations were rated as being highly semantically similar across languages. At least 25{\%} of the words in each direction of translation had more than one translation. The form similarity ratings were found to be highly reliable even when obtained with different bilinguals and modified rating procedures. Number of translations and meaning factors significantly predicted the semantic similarity of translation pairs. In future research, these norms may be used to determine the number of translations of words to control for or study this factor. These norms are available at http://www.talkbank.org/norms/tokowicz/.
The aim of the present study was to provide Russian normative data for the Snodgrass and Vanderwart (Behavior Research Methods, Instruments, {\&} Computers, 28, 516-536, 1980) colorized pictures (Rossion {\&} Pourtois, Perception, 33, 217-236, 2004). The pictures were standardized on name agreement, image agreement, conceptual familiarity, imageability, and age of acquisition. Objective word frequency and objective visual complexity measures are also provided for the most common names associated with the pictures. Comparative analyses between our results and the norms obtained in other, similar studies are reported. The Russian norms may be downloaded from the Psychonomic Society supplemental archive.
Ratings of age of acquisition (AoA), imageability, and familiarity were collected for 1,526 words. The methodology made use of a modular approach, in which the full sample of words was divided into five separate blocks. Within each block, each word was rated on each of the three variables by 20 partici- pants (undergraduate students from the University of Bristol). Analyses comparing these ratings to existing norm databases demonstrated that this methodology resulted in high reliability (assessed by Cronbach's ) and validity. The ratings were also transformed to be compatible with the Gilhooly and Logie (1980) norms. This transformation resulted in a set of norms for 3,394 words, which is by far the largest database of ratings for AoA, imageability, and familiarity to date. The resulting database should be useful for researchers interested in manipulating or controlling these factors in word recognition, neuropsychological, or memory studies. These norms can be downloaded from language.psy.bris .ac.uk/bristol{\_}norms.html.
Orthographic transparency metrics for opaque or deep languages, such as French and English, have tended to focus on feedforward and/or feedback directions, with claims made for the influence of both on reading. In the present study, data for five transparency metrics for southern British English, three of which are neither feedforward nor feedback, are presented, demonstrating the complex relationships between the metrics and offering an explanation for feedback effects in children's reading accuracy. The structure of such metrics from a variety of corpus sizes and origins is investigated, and it is concluded that large corpus sizes do not make a substantial contribution to the value of such metrics, when compared with smaller samples, and that adult and child corpuses have very similar profiles. Probabilities of occurrence for the phonemes, graphemes, and sonographs in this study may be downloaded from brm.psychonomic-journals.org/content/supplemental.
Matching stimuli across a range of influencing variables is no less important for studies of face recognition than it is for those of word processing. Whereas a number of corpora exist to allow experimenters to select a carefully controlled set of word stimuli, similar databases for famous faces do not exist. This article, therefore, provides researchers in the area of face recognition with a useful resource on which to base their stimulus selection. In the first phase of the investigation, British adults over 40 years of age were requested to generate the names of famous people (or celebrities) that they thought they would recognize and to write these down. The most frequently named celebrities were then rated by adults from the same age population for familiarity, distinctiveness, and age of acquisition. The result is a database of 696 famous people, with an indication of their relative eminence in the public consciousness and rated for these important variables. Phoneme counts are also provided for each famous person, together with family name frequency counts in the general population, where available. Materials and links may be accessed at www.psychonomic.org/archive.
A three-phased study was conducted in order to develop a standardized list of touch-related adjec-tives. The final list consisted of 306 words that were categorized in 440 instances according to the Le-derman and Klatzky (1987, 1990)dimensions of haptic properties (some words were classified in more than one dimension). The Kucera and Francis (1967)frequency of occurrence in written English for all words in the final list was also determined. A correlation was found between frequency of occurrence on the list and Kucera and Francis frequency. An analysis of the word dimensions and future applica-tions are discussed.
Planning, predicting, reasoning, and acting often depend crucially on the correct encoding and application of knowledge concerning the temporal and causal ordering of events. Yet no pictorial stimulus set is optimized for investigating the processing of temporal and causal order information. We introduce a novel stimulus set of 265 black-and-white line drawings depicting a diverse array of recognizable events. Most of the images in the stimulus set (N = 222) share a thematic or conceptual association with one other image in the set, and the stimuli were created and extensively normed such that the image pairs vary in the degrees to which they share a causal, ordered relation with one another. The stimuli were standardized in a series of normative tasks, including concept/noun/verb agreement, perceived frequency, visual similarity, and indexes of three features of causal associations between events (i.e., temporal proximity, exclusivity, and priority). Both younger adults (ages 18-30 years) and older adults (ages 60-80 years) contributed normative data, allowing for broad applications of the stimuli to the study of normal and age-related changes in the encoding, retention, and retrieval of information regarding temporal and causal order. Complete normative data sets are available in the online supplemental materials, and the full stimulus set is available by contacting the first author.
In this article, we describe the most extensive set of word associations collected to date. The database contains over 12,000 cue words for which more than 70,000 participants generated three responses in a multiple-response free association task. The goal of this study was (1) to create a semantic network that covers a large part of the human lexicon, (2) to investigate the implications of a multiple-response procedure by deriving a weighted directed network, and (3) to show how measures of centrality and relatedness derived from this network predict both lexical access in a lexical decision task and semantic relatedness in similarity judgment tasks. First, our results show that the multiple-response procedure results in a more heterogeneous set of responses, which lead to better predictions of lexical access and semantic relatedness than do single-response procedures. Second, the directed nature of the network leads to a decomposition of centrality that primarily depends on the number of incoming links or in-degree of each node, rather than its set size or number of outgoing links. Both studies indicate that adequate representation formats and sufficiently rich data derived from word associations represent a valuable type of information in both lexical and semantic processing.
Age of acquisition (AoA) is an important psycholinguistic variable that affects the speed and accuracy of lexical processing in tasks such as word naming, picture naming, and lexical decision. In the present work, we collected AoA ratings for 1,749 Portuguese words (nouns, verbs, adjectives, and adverbs), using a 9-point scale that was first proposed by Carroll and White (1973). We analyzed the relation between AoA ratings and other psycholinguistic variables (length measures, neighborhood density, written-word frequency, familiarity, imageability, and concreteness), and we assessed reliability by correlating our ratings with those from other databases presented for Portuguese, English, Spanish, and Italian. The full database can be downloaded from http://brm.psychonomic-journals.org/content/supplemental.
Two experiments attempted to resolve previous contradictory findings concerning developmental trends in false memories within the Deese-Roediger-McDermott (DRM) paradigm by using an improved methodology--constructing age-appropriate associative lists. The research also extended the DRM paradigm to preschoolers. Experiment 1 (N=320) included children in three age groups (preschoolers of 3-4 years, second-graders of 7-8 years, and preadolescents of 11-12 years) and adults, and Experiment 2 (N=64) examined preschoolers and preadolescents. Age-appropriate lists increased false recall. Although preschoolers had fewer false memories than the other age groups, they showed considerable levels of false recall when tested with age-appropriate materials. Results were discussed in terms of fuzzy-trace, source-monitoring, and activation frameworks.
Used correlation functions obtained in 2 experiments with undergraduate Os (N = 10) as a basis for describing human visual letter recognition. Visual images were filtered by means of autocorrelation for pattern information. This operation gave the relative visibilities or legibilities of the characters. The visual impressions were then cross-correlated with a set of memory records whose outputs described the relative probabilities that the stimulus was a given character. This operation described confusion errors. Finally "response bias" was described in terms of the reliability with which a memory record provides identification of a given stimulus. In these terms response bias represented an attempt by the recognition system to minimize errors in high-information responses, at the expense of producing more low-information responses as errors. (French summary)
This article introduces EsPal: a Web-accessible repository containing a comprehensive set of properties of Spanish words. EsPal is based on an extensible set of data sources, beginning with a 300 million token written database and a 460 million token subtitle database. Properties available include word frequency, orthographic structure and neighborhoods, phonological structure and neighborhoods, and subjective ratings such as imageability. Subword structure properties are also available in terms of bigrams and trigrams, biphones, and bisyllables. Lemma and part-of-speech information and their corresponding frequencies are also indexed. The website enables users either to upload a set of words to receive their properties or to receive a set of words matching constraints on the properties. The properties themselves are easily extensible and will be added over time as they become available. It is freely available from the following website: http://www.bcbl.eu/databases/espal/ .
Individual happiness is a fundamental societal metric. Normally measured through self-report, happiness has often been indirectly characterized and overshadowed by more readily quantifiable economic indicators such as gross domestic product. Here, we examine expressions made on the online, global microblog and social networking service Twitter, uncovering and explaining temporal variations in happiness and information levels over timescales ranging from hours to years. Our data set comprises over 46 billion words contained in nearly 4.6 billion expressions posted over a 33 month span by over 63 million unique users. In measuring happiness, we use a real-time, remote-sensing, non-invasive, text-based approach---a kind of hedonometer. In building our metric, made available with this paper, we conducted a survey to obtain happiness evaluations of over 10,000 individual words, representing a tenfold size improvement over similar existing word sets. Rather than being ad hoc, our word list is chosen solely by frequency of usage and we show how a highly robust metric can be constructed and defended.
This article describes a Windows program that enables users to obtain a broad range of statistics concerning the properties of word and nonword stimuli in Spanish, including word frequency, syllable frequency, bigram and biphone frequency, orthographic similarity, orthographic and phonological structure, concreteness, familiarity, imageability, valence, arousal, and age-of-acquisition measures. It is designed for use by researchers in psycholinguistics, particularly those concerned with recognition of isolated words. The program computes measures of orthographic similarity online, with respect to either a default vocabulary of 31,491 Spanish words or a vocabulary specified by the user. In addition to providing standard orthographic and phonological neighborhood measures, the program can be used to obtain information about other forms of orthographic similarity, such as transposed-letter similarity and embedded-word similarity. It is available, free of charge, from the following Web site: www.maccs.mq.edu.au/-colin/B-Pal.
In this article, normative data on the familiarity and difficulty of 196 single-solution Spanish word fragments are presented. The database includes the following indices: difficulty, familiarity, frequency, number of meanings, number of letters given in the fragment, first and/or last letters given, and ratio of letters to blanks. A factor analysis was performed on difficulty, and two factors were obtained. Frequency, familiarity, and number of meanings loaded highly on the first factor, which we consider to measure lexical processes, whereas number of letters in the fragment, first and/or last letters given, and ratio of letters to blanks loaded highly on the second factor, which we judge to be determined by perceptual information. Regression analyses using factor scores as predictors showed that both factors accounted for a significant part of the completion probability scores. The full set of these norms may be downloaded from the Psychonomic Society Web archive at