1358 norm sets
CONTAINS 3 LISTS FOR USE IN ANAGRAM STUDIES: (1) A SOURCE LIST GROUPING WORDS ACCORDING TO THEIR ALPHABETIZED FORM, (2) ALL 5-LETTER ENGLISH WORDS WHOSE LETTERS CANNOT BE REARRANGED TO SPELL ANY OTHER WORD, AND (3) AN ALPHABETICAL LISTING OF ALL MULTIPLE- AND SINGLE-SOLUTION 5-LETTER WORDS. (PsycINFO Database Record (c) 2016 APA, all rights reserved)
NORMATIVE SOLUTION TIMES BASED ON A SAMPLE OF 134 SOLUTION WORDS AND 378 ASSOCIATED ANAGRAMS COMPILED FROM 9 STUDIES ARE PRESENTED, AS WELL AS THE 120 LETTER ORDERS POSSIBLE WITH A 5-LETTER WORD, AND A SKELETON-WORD TEST AND SCORING KEY USED FOR ASSESSING THE DEGREE TO WHICH SS STORE DIGRAM FREQUENCY INFORMATION. (16 REF.)
PRESENTS TABLES, BASED ON A SAMPLE OF 20,000 ENGLISH WORDS, WHICH SHOW SINGLE-LETTER AND DIGRAM FREQUENCY COUNTS BROKEN DOWN TO ACCOUNT FOR ALL WORD-LENGTH AND LETTER-POSITION COMBINATIONS, FOR WORDS 3-7 LETTERS IN LENGTH. THESE TABLES ARE SIMILAR TO THOSE OF PRATT AND OF UNDERWOOD AND SCHULZ, BUT IN ADDITION ALLOW FOR THE ASSESSMENT OF FREQUENCIES FOR ALL WORD-LENGTH AND LETTER-POSITION VARIATIONS IN THE WORD SAMPLE. THE TABLES MAY BE EMPLOYED TO PROVIDE NORMATIVE FREQUENCY DATA FOR STUDIES IN VERBAL LEARNING AND RETENTION, ANAGRAM PROBLEM SOLVING, WORD RECOGNITION THRESHOLDS, LINGUISTIC ANALYSES, ETC. (PsycINFO Database Record (c) 2012 APA, all rights reserved)
A WORD LIST OF SPOKEN RUSSIAN WAS COMPILED BASED ON AN ACTUAL COUNT OF 10,000 WORDS. THE WORDS WERE COMPOSED OF 50-WORD SAMPLES TAKEN FROM 200 ACTS OF 93 PLAYS PUBLISHED SINCE 1957. IT WAS FOUND THAT JUST 360 WORDS, FROM A TOTAL OF 2,380 WORDS TABULATED, REPRESENTED 73 PERCENT OF ALL OCCURRENCES. THE AUTHOR PREPARED SAMPLE DIALOGUES USING ONLY THESE HIGH-FREQUENCY VOCABULARY ITEMS AS FURTHER PROOF THAT INTELLIGENT COMMUNICATION AT AN ADULT LEVEL WAS POSSIBLE.
Word association norms are presented for 2 lists of 200 words including 100 words of the Kent-Rosanoff list. Ss were 250 boys and 250 girls in each of the Grades 4-8, 10, and 12 in the Minneapolis public schools and 500 male and 500 female students in introductory psychology classes at the University of Minnesota. (PsycINFO Database Record (c) 2007 APA )
A re-evaluation of all possible 3-letter combinations of the Roman alphabet of the form consonant-vowel-consonant with the restriction that the 2 consonants are different and that neither is a y when y is the vowel. A list of 2480 possible trigrams is included. Reliability was determined by 2 measures, and the association values were correlated with the previously reported list of trigrams by Glaze and Krueger. It was concluded that the present list provides a better estimate of the meaningfulness of trigrams than has thus far been available. From Psyc Abstracts 36:01:1CI23A. (PsycINFO Database Record (c) 2006 APA, all rights reserved)
In this work we present SENTIWORDNET 3.0, a lexical resource explicitly devised for supporting sentiment classification and opinion mining applications. SENTIWORDNET 3.0 is an improved version of SENTIWORDNET 1.0, a lexical resource publicly available for research purposes, now currently licensed to more than 300 research groups and used in a variety of research projects worldwide. Both SENTIWORDNET 1.0 and 3.0 are the result of automatically annotating all WORDNET synsets according to their degrees of positivity, negativity, and neutrality. SENTIWORDNET 1.0 and 3.0 differ (a) in the versions of WORDNET which they annotate (WORDNET 2.0 and 3.0, respectively), (b) in the algorithm used for automatically annotating WORDNET, which now includes (additionally to the previous semi-supervised learning step) a random-walk step for refining the scores. We here discuss SENTIWORDNET 3.0, especially focussing on the improvements concerning aspect (b) that it embodies with respect to version 1.0. We also report the results of evaluating SENTIWORDNET 3.0 against a fragment of WORDNET 3.0 manually annotated for positivity, negativity, and neutrality; these results indicate accuracy improvements of about 20{\%} with respect to SENTIWORDNET 1.0.
India is a multilingual country where machine translation and cross lingual search are highly relevant problems. These problems require large resources- like wordnets and lexicons- of high quality and coverage. Wordnets are lexical structures composed of synsets and semantic relations. Synsets are sets of synonyms. They are linked by semantic relations like hypernymy (is-a), meronymy (part-of), troponymy (manner-of) etc. IndoWordnet is a linked structure of wordnets of major Indian languages from Indo-Aryan, Dravidian and Sino-Tibetan families. These wordnets have been created by following the expansion approach from Hindi wordnet which was made available free for research in 2006. Since then a number of Indian languages have been creating their wordnets. In this paper we discuss the methodology, coverage, important considerations and multifarious benefits of IndoWordnet. Case studies are provided for Marathi, Sanskrit, Bodo and Telugu, to bring out the basic methodology of and challenges involved in the expansion approach. The guidelines the lexicographers follow for wordnet construction are enumerated. The difference between IndoWordnet and EuroWordnet also is discussed.
Die vorliegende Dissertation beschreibt Aufbau und Funktionalit{\"{a}}t der bayerischen Dialektdatenbank BAYDAT. Die Datenbank fasst die Erhebungsdaten der Teilprojekte des Bayerischen Sprachatlas (BSA) zusammen, speichert sie zukunftssicher und macht sie zentral nutzbar. Die Arbeit zeigt die Vorgehensweise bei der Aufbereitung der Quelldateien, beschreibt die einzelnen Datenbanktabellen der BAYDAT-Datenbank und widmet sich der Realisierung und Funktionalit{\"{a}}t der Onlineoberfl{\"{a}}che, {\"{u}}ber die die BAYDAT-Datenbank einem weltweiten Nutzerkreis aus Dialektologen und interessierten Laien zur Verf{\"{u}}gung stehen soll. The PhD thesis at hand describes the setup and functionality of the Bavarian dialect database BAYDAT. The database integrates the data of the subprojects of the Bayerischer Sprachatlas (BSA, Atlas of the Bavarian language). The database ensures a future-proof storage of the data and makes the data available at one central point. The thesis describes the methods used in formatting the source data. It also describes the structure of the different database tables. It also contains a description of the development and the functionality of BAYDAT's graphical user interface that will allow linguists and laypersons to access the database online.
Today, people generate and store more data thanever before as they interact with both real and virtual environ-ments. These digital traces of behavior and cognition offercognitive scientists and psychologists an unprecedented op-portunity to test theories outside the laboratory. Despite gen-eral excitement about big data and naturally occurring datasetsamong researchers, threeBgaps{\^{}}stand in the way of theirwider adoption in theory-driven research: theimaginationgap, theskillsgap, and theculturegap. We outline an ap-proach to bridging these three gaps while respecting our re-sponsibilities to the public as participants in and consumers ofthe resulting research. To that end, we introduce Data on theMind (http://www.dataonthemind.org), a community-focusedinitiative aimed at meeting the unprecedented challenges andopportunities of theory-driven research with big data and nat-urally occurring datasets. We argue that big data and naturallyoccurring datasets are most powerfully used tosupplement—not supplant—traditional experimental para-digms in order to understand human behavior and cognition,and we highlight emerging ethical issues related to the collec-tion, sharing, and use of these powerful datasets.
Languages with binary stress systems frequently tolerate a stress lapse over the final two syllables, but almost none tolerate a word-initial stress lapse. Lunden (to appear) argues that this lapse asymmetry can be explained by the presence of word-level final lengthening, which can then create the perception of prominence alternation in languages that use duration as stress correlate. The results of a production and a perception study with English speakers are presented which compare /ɑ/s that occur under stress lapse to /ɑ/s in non-stress-lapse positions. While word-final unstressed /ɑ/ is always longer than non-final unstressed /ɑ/, it is significantly longer when immediately following an unstressed syllable. Similarly, unstressed word-final /ɑ/ has a higher F1 and lower F2 than non-final unstressed /ɑ/, but word-finally this less-reduced vowel is closer to a full vowel when the final syllable is part of a stress lapse. The perception study finds that these differences have perceptual consequences that can lead to a perceived continued rhythm in stress lapse. The phonetic differences explain why a word-final unstressed vowel can be perceived as relatively strong when following an unstressed syllable but as relatively weak when following a stressed syllable.
TLS explores the conceptual schemes of pre-Buddhist Chinese on the basis of over 8500 A4 pages of text with interlinear translations. TLS is a sustained effort in philological and philosophical fieldwork, designed throughout to make the classical Chinese evidence strictly comparable to that of other cultures, and to make possible meaningful analytic primary-evidence-based disagreement among non-sinologists on classical Chinese concepts and words. TLS is compiled in the hope that careful philosophical reflection on Chinese texts might serve to broaden the empirical basis for philosophical theories and generalisations on conceptual schemes. TLS is based on the conviction that we should improve the clarity and bite of declarations of difference between conceptual schemes by enlarging the basis of literally translated and analysed texts from widely (though never radically) different intellectual cultures. The necessary charitable assumption that if we want to understand others we must count them right in most matters will not prevent TLS from looking for and exploring deep conceptual contrasts to the full. TLS seeks to make precise criteria of translation for classical Chinese, mainly through a detailed description in English of systematic recurrent semantic relations between Chinese words, especially distinctive semantic features. TLS is the first synonym dictionary of classical Chinese in any Western language. TLS focusses on distinctive semantic nuances. TLS is the first interactive dictionary of Chinese. TLS is the first dictionary which systematically organises the Chinese vocabulary in taxonomic and mereonomic hierarchies thus showing up whole conceptual schemes or cognitive systems. These are taken to circumscribe the changing topology of Chinese mental space. TLS is the first dictionary that systematically registers a range of lexical relations like antonym, converse, epithet etc. TLS thus aims to define conceptual space as a relational space. TLS is the first dictionary of Chinese which incorporates detailed syntactic analysis of (over 600 distinct kinds of) syntactic usage. TLS thus enables us to make a systematic study of such basic phenomena as the natural history of abstract nouns in China. TLS is the first corpus-based dictionary which will record the history of rhetorical devices in texts and will thus enable us to study such intellectually crucial things as the natural history of irony in China. All analytic categories and procedures of analysis in TLS are flexible in the sense that they are continuously being revised and improved in the light of new observation and analysis
This book promotes the development of linguistic databases by describing a number of successful database projects, focusing especially on cross-linguistic and typological research. It has become increasingly clear that ready access to knowledge about cross-linguistic variation is of great value to many types of linguistic research. Such a systematic body of data is essential in order to gain a proper understanding of what is truly universal in language and what is determined by specific cultural settings. Moreover, it is increasingly needed as a tool to systematically evaluate contrasting theoretical claims. The book includes a chapter on general problems of using databases to handle language data and chapters on a number of individual projects.
This landmark publication in comparative linguistics is the first comprehensive work to address the general issue of what kinds of words tend to be borrowed from other languages. The authors have assembled a unique database of over 70,000 words from 40 languages from around the world, 18,000 of which are loanwords. This database (http://loanwords.info) allows the authors to make empirically founded generalizations about general tendencies of word exchange among languages"--Provided by publisher. Notational conventions -- Acknowledgments -- List of authors -- General chapters: I. The loanword typology project and the world loanword database / Martin Haspelmath and Uri Tadmor -- II. Lexical borrowing: Concepts and issues / Martin Haspelmath -- III. Loanwords in the world's languages: Findings and results / Uri Tadmor -- THE LANGUAGES: 1. Loanwords in Swahili / Thilo C. Schadeberg -- 2. Loanwords in Iraqw, a Cushitic language of Tanzania / Maarten Mous and Martha Qorro -- 3. Loanwords in Gawwada, a Cushitic language of Ethiopia / Mauro Tosco -- 4. Loanwords in Hausa, a Chadic language in West Africa / Ari Awagana and H. Ekkehard Wolff, with Doris Löhr -- 5. Loanwords in Kanuri, a Saharan language / Doris Löhr and H. Ekkehard Wolff, with Ari Awagana -- 6. Loanwords in Tarifiyt, a Berber language of Morocco / Maarten Kossmann -- 7. Loanwords in Seychelles Creole / Susanne Michaelis with Marcel Rosalie -- 8. Loanwords in Romanian / Kim Schulte -- 9. Loanwords in Selice Romani, an Indo-Aryan language of Slovakia / Viktor Elšík -- 10. Loanwords in Lower Sorbian, a Slavic language of Germany / Hauke Bartels -- 11. Loanwords in Old High German / Roland Schuhmann -- 12. Loanwords in Dutch / Nicoline van der Sijs -- 13. Loanwords in British English / Anthony Grant -- 14. Loanwords in Kildin Saami, a Uralic language of northern Europe / Michael Riessler -- 15. Loanwords in Bezhta, a Nakh-Daghestanian of the North Caucasus / Bernard Comrie and Madzhid Khalilov -- 16. Loanwords in Archi, a Nakh-Daghestanian of the North Caucasus / Marina Chumakina -- 17. Loanwords in Manange, a Tibeto-Burman language of Nepal / Kristine A. Hildebrandt -- 18. Loanwords in Ket, a Yeniseian language of Siberia / Edward Vajda -- 19. Loanwords in Sakha (Yakut), a Turkic language of Siberia / Brigitte Pakendorf and Innokentij N. Novgorodov -- 20. Loanwords in Oroqen, a Tungusic language of China / Fengxiang Li and Lindsay J. Whaley. Loanwords in Japanese / Christopher K. Schmidt -- 22. Loanwords in Mandarin Chinese / Thekla Wiebusch and Uri Tadmor -- 23. Loanwords in Thai / Titima Suthiwan and Uri Tadmor -- 24. Loanwords in Vietnamese / Mark J. Alves -- 25. Loanwords in White Hmong / Martha Ratliff -- 26. Loanwords in Ceq Wong, an Austroasiatic language of Peninsular Malaysia / Nicole Kruspe -- 27. Loanwords in Indonesian / Uri Tadmor -- 28. Loanwords in Malagasy / Alexander Adelaar -- 29. Loanwords in Takia, an Oceanic language of Papua New Guinea / Malcolm Ross -- 30. Loanwords in Hawaiian / 'Ōiwi Parker Jones -- 31. Loanwords in Gurindji, a Pama-Nyungan language of Australia / Patrick McConvell -- 32. Loanwords in Yaqui, a Uto-Aztecan language of Mexico / Zarina Estrada Fernández -- 33. Loanwords in Zinacantán Tzotzil, a Mayan language of Mexico / Cecil H. Brown -- 34. Loanwords in Q'eqchi', a Mayan language of Guatemala / S{\o}ren Wichmann and Kerry Hull -- 35. Loanwords in Otomi, an Otomanguean language of Mexico / Ewald Hekking and Dik Bakker -- 36. Loanwords in Saramaccan, an English-based creole of Suriname / Jeff Good -- 37. Loanwords in Imbabura Quechua / Jorge Gómez Rendón and Willem Adelaar -- 38. Loanwords in Kali'na, a Cariban language of French Guiana / Odile Renault-Lescure -- 39. Loanwords in Hup, a Nadahup language of Amazonia / Patience Epps -- 40. Loanwords in Wichí, a Mataco-Mataguayan language of Argentina / Alejandra Vidal and Verónica Nercesian -- 41. Loanwords in Mapudungun, a language of Chile and Argentina / Lucía A. Golluscio.