Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
16504 papers
This research assessed an interactive satellite-based training program integrating interactive audiovisual experiences with face-to-face interactions. Key elements were content created by experts, high-quality video segments, satellite-based interaction, off-line interactions among teams of parents and caregivers, workshops, and team building exercises. For pragmatic reasons, it was necessary to develop brief assessment instruments concurrently with training. A large set of survey items were created from draft materials and reduced empirically through piloting to those with the best psychometric properties. To avoid the appearance of traditional testing, knowledge was assessed with Likert items. Surveys measured participant satisfaction, knowledge, attitudes, and the application and articulation of concepts. Participant satisfaction was high. Participants increased positive attitudes and learned appropriate vocabulary. Training was more effective than no training or watching videotapes. The program appears to represent a viable model of training that could successfully be applied to Internet technologies.
This paper focuses on ‘English’ dictionaries and their development in second language learning contexts, taking the perspective that ‘standards’ are usually codified in reference grammars, pronouncing dictionaries and word dictionaries. It begins with a presentation of contemporary discussions of ‘English’ and ‘Englishes’ in Asia, a phenomenon that has come about through the global spread of what is now a truly universal language. With the second diaspora of English (Kachru, 1992), many of the educational institutions in Asia are beginning to feel the tension between rigid and loose canons, and between traditional and emerging norms of language usage. In English as an additional language learning communities, lexical innovation is evident in new canons, and in primary sources of data such as newspapers. How do dictionaries handle such innovations, and become themselves a secondary source? The question may not be an easy one to answer but the discussion leading to it may herald a new dawn in dictionary‐making in Asia.
The FUL (featurally underspecified lexicon) model of automatic speech recognition is based on the representation of words in the lexicon with underspecified distinctive features. The speech signal is converted from the waveform into an online spectral representation made up of formants and a few parameters describing the overall spectral shape. These LPC and spectral parameters are converted into distinctive phonological features which, in turn, are compared with all entries in the lexicon. No classification into segments, syllables, or spectral templates is used for the selection of words from the lexicon. Comparison of signal features with those stored in the lexicon uses a ternary system of matching, nomismatching, and mismatching features. Matching features increase the scoring for potential word candidates, no-mismatching features do not exclude candidates and only mismatching features lead to the rejection of word candidates. The word candidates are expanded to include word hypotheses, even without further acoustic evidence, and are used in the phonological and syntactic parsing that operates in parallel with the acoustic front-end. 1. THEORY The speech signal of the same phonetic segment varies across dialects and speakers — within speakers in certain segmental and prosodic contexts, and even for the same speaker and context with repetition, speaking rate, emotional state, microphones, etc. Not surprisingly, speech recognition with simple spectral template matching has failed consistently. Any variation in the signal leads to variation of the spectra that are compared to the stored templates. Only statistical approaches like Hidden Markov Models based on large training sets have led to acceptable results, but are still speaker and transmission-line dependent or operate only with a restricted vocabulary, syntax, and semantics. Human listeners seem to be unconcerned by adverse acoustical conditions and are able to resolve a wide range of variations like assimilations and deletions with apparent ease. Ambiguities in the signal, whether they come from random noise or whether they are linguistic in nature, like cliticizations of words or assimilations are the norm rather than the exception in natural language. Human listeners, however, appear not to be worried by adverse acoustic conditions and indeed, handle “variations” in the signal with ease. Language comprehension experiments [1, 2] have shown that listeners extract certain acoustic characteristics reliably but do not match acoustic details with the lexicon. Rather, the experimental results are best explained with the assumption that lexical access involves mapping the acoustic signal to an underspecified featural representation. For example, the assimilation of a coronal sound (e.g. /n/) to a following labial place of articulation (like [b] in “Where could Mr. Bean be?”) often results in the production of a labial (i.e. “Bea[m] be”). The reverse is not true, that is, a labial sound does not assimilate to a coronal place of articulation (i.e., “la[m]e duck” does not become “la[n]e duck”). Simple articulatory mechanics cannot account for such behaviour because an articulatory assimilation would operate in both directions. An explanation can be given by assuming that coronal sounds are underspecified for place, whereas labial and dorsals are not: the labial place of articulation spreads to the preceding coronal sound (if the language has regressive assimilation) because that sound is not specified for place. On the other hand, the specification of a labial place prevents the place features of an adjacent sound from overriding this information. Consequently, coronal sounds can become labial (or dorsal), but labial (or dorsal) sound cannot change their place. This explanation is straightforward for speech production, but what about speech perception? How can a realisation of “gree[m]” in a labial context (like “bag”) or “gree[N]” in a dorsal context (like “grass”) lead to the access of the word “green” in the lexicon? Normally, “gree[m]” and “gree[N]” are nonwords in English. And, at the same time, how should a mechanism be constructed to allow the activation of the word “bean” as well as “beam” if the acoustic input is “bea[m]”, when “bean” is a word of the language? Human listeners handle these asymmetries (and many other assimilatory effects) within and across words without noticing it, as reaction-time experiments have shown [4]. The solution to these seemingly contradictory requirements can be obtained (i) by assuming an underspecified representation in the lexicon, where certain features (like the place feature [coronal]) are not stored in the lexicon (in speech production, segments with unspecified place are generated with the feature coronal by default) and (ii) by postulating a ternary matching logic in the signal-to-lexical mapping. page 715 ICPhS99 San Francisco candidate 1 candidate 2 • • • candidate n 1
This paper provides normative data for Australian school children on a modified version of the Castles Word/Non-Word Tests (Castles, 1993). The tests were designed to isolate the lexical and nonlexical reading procedures. Data were collected from 298 school children in Perth and combined with data provided by Coltheart and Leahy (1996) for 420 school children in Sydney. Norms for the Sydney sample have been published previously (Coltheart & Leahy, 1996). Norms for the combined sample are reported in 12-month age bands from 7 to 12 years, in the form of normalised standard scores. Issues surrounding subtyping research in dyslexia are reviewed, and a way to subclassify research samples using the provided norms is outlined and evaluated.
The journals of the Psychonomic Society have served as outlets for numerous stimulus norms and ratings. Such norms are useful to researchers in a variety of areas for manipulating and controlling stimulus attributes. This article presents an index of 142 norms published in the Society’s journals, categorized according to the types of materials and ratings that are included in each.
This paper describes morphing techniques to manipulate two-dimensional human face images and three-dimensional models of the human head. Applications of these techniques show how to generate composite faces based on any number of component faces, how to change only local aspects of a face, and how to generate caricatures and anticaricatures of faces. These techniques are potentially useful for many psychological studies because they permit realistic images to be generated with precise control.
Electronic texts are claimed to exhibit features distinct from their more tangible cousins. The Snapshot project aims to observe and capture language usage in an electronic medium by creating an open corpus of World Wide Web documents. These documents are re-encoded using the TEI guidelines to create a flexible, persistent and portable data repository. This report gives an overview of the decisions made with respect to the re-encoding of HTML documents, and with the structuring the overall corpus.
Boundary extension refers to a tendency to remember seeing a greater expanse of a scene than was shown in a photograph. It is hypothesized that the view shown in the stimulus activates expectations about the scene’s layout just outside the picture’s borders. Following presentation, the viewer remembers having seen this expected information, and this yields boundary extension. We provide photographs and instructions for conducting two brief demonstrations of the phenomenon and provide materials for a related class experiment on the journal’s World-Wide Web site. These demonstrations of boundary extension provide graphic illustrations of the role of schematic expectancies in the representation of scenes and help to illustrate the role of real-world knowledge in cognition.
A number of studies in perception, attention, and memory employ signal detection theory (SDT) to assess the accuracy of an observer’s detection or discrimination performance. Some of the problems that students have with understanding and using SDT are associated with the calculations needed to obtain SDT parameters and predictions. All of these calculations, plus the simulation of SDT processes, can be performed using a spreadsheet application program, such as Excel or Quattro Pro. This paper offers a short tutorial on how to use a spreadsheet program to increase your students’ knowledge and understanding of SDT.
The Internet presents a potentially revolutionary tool in the dissemination of scientific information, offering many advantages to authors and audiences. However, this resource has been underutilized in psychological research because of several factors: unfamiliarity with required technology, lack of peer review, absence of an efficient centralized accessibility resource, concerns about copyright issues, and financial considerations. The present article describes the advantages of on-line presentation of research, as well as discusses various concerns about on-line publishing and the developing solutions to deal with those concerns.
This paper explores the role of lexicalization and prun-ing of grammars for base noun phrase identification. We modify our original framework (Cardie & Pierce 1998) to extract lexicalized treebank grammars that assign a score to each potential noun phrase based upon both the part-of-speech tag sequence and the word sequence of the phrase. We evaluate the mod-ified framework on the “simple ” and “complex ” base NP corpora of the original study. As expected, we find that lexicalization dramatically improves the perfor-mance of the unpruned treebank grammars; however, for the simple base noun phrase data set, the lexical-ized grammar performs below the corresponding unlex-icalized but pruned grammar, suggesting that lexical-ization is not critical for recognizing very simple, rel-atively unambiguous constituents. Somewhat surpris-ingly, we also find that error-driven pruning improves the performance of the probabilistic, lexicalized base noun phrase grammars by up to 1.0 % recall and 0.4% precision, and does so even using the original pruning strategy that fails to distinguish the effects of lexical-ization. This result may have implications for many probabilistic grammar-based approaches to problems in natural language processing: error-driven pruning is a remarkably robust method for improving the perfor-mance of probabilistic and non-probabilistic grammars alike.
This paper describes a compound unit (CU) recognizer as a pattern‐based approach and its hybridization with rule‐based translation. A compound unit is a combined concept including collocations, idioms, and compound nouns. CU recognition reduces part of speech ambiguities by combining several words into a unit and consequently lessening the parsing load. It also provides pretranslated natural equivalents. Our focus in this paper is to obtain flexibility and efficiency from pattern‐based machine translation, and high‐quality translation by hybridization. A modified trie, our search index structure using “method” strategy is used to manage heterogeneous property of the constituents. Syntactic verification is integrated to obtain precise CU recognition by means of pruning wrongly recognized units that are caused by improper variable hypotheses. The experimental result with verification shows that the precision of CU recognition is increased to 99.69% with 31 CFG rules on the cyclic trie structure for 1,268 Wall Street Journal articles of the Penn Treebank. Another experiment with CU recognition also shows that it raises the understandability of translation for Web documents.
This study describes the typical course and variability in major areas of communicative development for 228 Swedish-speaking children between 8 and 16 months of age. The assessments were made by parental reports with the Swedish Early Communicative Development Inventories (SECDI) using a semi-longitudinal design. Age-based norms for understanding of phrases, vocabulary comprehension, vocabulary production and use of gestures are described at the 10th, 25th, 50th, 75th and 90th percentile levels. More lexical verbs were found among the first words in comprehension than in production. An extensive variability within individuals in onset and development was found for the assessed skills. The individual differences proved to be stable over 4–6 months. No gender differences were found for comprehension of phrases, total gestures, vocabulary compre-hension, or for vocabulary production. Strong, unique associations were found between total gestures and vocabulary comprehension and between vocabulary comprehension and vocabulary production. In contrast, no unique association was found between gestures and vocabulary production. The results generally concur with those reported for English-speaking American children by Fenson et al. (1993, 1994).
Research has shown that seeing another person (i.e., a model) perform at a certain level can influence the goal choice and performance of an observer. This study extended these findings by examining the model's effects on two potential explanatory mediators: expectations of reaching different performance levels and valence at those levels. The moderating effects of observer task experience and self‐esteem were also examined. In a repeated measures design, results showed that model performance influenced observers' goals, task performance, expectancies, and valence ratings. Path analyses indicated that expectancies mediated observational effects on goal choice. Results also indicated significant Model x Trial interactions, with a diminishing effect of model influence as personal experience developed. Results are discussed in terms of mediating effects of expectancies and valences on model influences on goal choice as well as the different weights given to social and personal information.
The traditions of grammatical terminology are quite different in the German and the French educational system. While the French ministry of education issues official regulations which are binding on a national level (they last did so in 1997), the federal system in Germany prevents the standardisation of the grammatical nomenclature for all the Länder of the FRG. Moreover, the historico-cultural background is different in both countries: the French public has been used to an active and centralist language policy for several centuries; in Germany governmental interference with the linguistic norm is often met with resistance, and the regulation of many details is in fact left to the important publishing houses. For this very reason it was left largely to the school-book publishing houses to decide how to put the recommendations of the German Secretaries of cultural affairs issued in 1982 into practice. The present article is based on a number of selected examples and provides a critical analysis of the usage of grammatical terminology in German and French school-books, and it pleads for a re-orientation – particularly in Germany: the inconsistencies and ad hoc solutions which can be observed quite frequently can only be overcome if the terminology is based on a well reflected linguistic theory. This is the only way to achieve an effect of synergy between grammar lessons in different school languages and ultimately a standardisation of the grammatical terminology on a European level.
Reversing a One-Way Bilingual Dictionary* Leonard Newmark "One day we will go back to Kosov[a]. That's our land." — Ramada Shaqiri, 30 March 1999 Afi fter completing a ten-year project to write an Albanian dicLtionary, published in March 1998 by Oxford University Press as the Oxford Albanian-English Dictionary and designed specifically for users who want to read Albanian and whose access language is English, I decided to prepare a companion English-Albanian dictionary for users who want to write Albanian, but I did not want to devote another ten years to that compilation. I wondered whether I could produce a useful bilingual dictionary with the reverse orientation by automatic conversion of the entries in the data files from which the first dictionary was generated. This paper is a report on the degree to which the attempt succeeded and the degree to which human intervention was required. Examples are provided to illustrate some rather surprising results, and a general conclusion is drawn for bilingual lexicography. For languages of limited worldwide commercial importance, like Albanian, it seems particularly important to use computational techniques to derive new dictionaries from lexical data files compiled for some other purpose, especially if those files are extensive and have information otherwise difficult to come by in machine-readable form. The richness of the lexical data files from which my Albanian-English dictionary was generated is evidenced by that dictionary's 75,000 entries and subentries, more than are found in any other dictionary of Albanian. Those files already provide a number of features that dis- *This paper is a reworked and expanded version of the paper I presented at the 8th EURALEX International Congress (4-8 August 1998) and published in the proceedings of that congress. 38Leonard Newmark tinguished this dictionary from many other bilingual dictionaries: 1) inclusion of large numbers of nonstandard items (marked by asterisks ) as well as all attested standard stems; 2) marking of morpheme boundaries in Albanian words; 3) inclusion of some 16,700 phrasal expressions, in particular, phrasal names, collocations, idioms, and proverbs; 4) use of large numbers of bipartite definitions with a discursive description of the sense followed, after a colon, by English synonyms exemplifying that sense; 5) inclusion of large numbers of terms for grasses, flowers, birds, and fish with their scientific definitions; 6) inclusion of a modest amount of encyclopedic information to explain words whose strictly lexical meaning would not make their use in Albanian contexts intelligible; 7) listing of the various stem forms of lexemes as separate entries in their own alphabetical position to enable readers to decipher otherwise mystifying forms encountered in actual texts; 8) indication of the specific limits of variation that leave idiomatic senses intact (e.g., in phrasal expressions, marking a verb that can appear in any of its inflected forms by giving it in citation form with a symbol (·) at the end of the stem); 9) elaborate labeling of Albanian distinctions in domain and register; 10) rendition of phrasal expressions by stylistically similar English expressions, frequently supplemented by literal translations (enclosed in quotation marks) to enable more nuanced understanding. Each entry in the plain text, flat data files from which the Albanian-English dictionary was generated is a line consisting of an Albanian word or phrase followed by a definition in English (or by a cross-reference to another line). Each line is embedded with simple visible two-letter formatting codes (e.g., HW [headword], DF [definition], TK [technical name], CO [collocation] ) immediately preceded by a period (.) and immediately followed by what will get the formatting assigned by that code. The easily redefinable codes are later translated by a set of UNIX scripts into formatting instructions in TeX, which can go directly to a printer or indirectly by translation into Post-Script files. The simplicity of such transparent and flexible coding for entering the data, in contrast with elaborate schemes requiring complex coding by experts into predefined structures,1 was initially dictated by limits typical of languages that attract little commercial interest and 'For example, those used in the architecture described by Willy Martin and Anne Tamm in "OMBI: An editor for constructing reversible lexical databases," EURALEX '96 Proceedings...
This paper explores the automatic construction of a multilingual Lexical Knowledge Base from pre-existing lexical resources. We present a new and robust approach for linking already existing lexical/semantic hierarchies. We used a constraint satisfaction algorithm (relaxation labeling) to select-among all the candidate translations proposed by a bilingual dictionary- the right English WordNet synset for each sense in a. taxonomy automatically derived from a Spanish monolingual dictionary. Although on average, there are 15 possible WordNet connections for each sense in the taxonomy, the method achieves an accuracy over 80%. Finally, we also propose several ways in which this technique could be applied to enrich and improve existing lexical databases.
This article discusses the principal problems of delimiting the scope of adverbs and their classification in terms of lexicology, lexicography and natural language processing, with a view to making the exhaustive lexical database for adverbs as well as to developing the description models which lead to the construction of the lexicon. Our discussion is not limited to the adverbs in traditional sense, but extended to the mono-/poly lexical words which can be considered as adverbials morphologically, syntactically and lexicologically. For instance, N + postposition(distributionally very limited), limited and/or productive inflectional forms of verbs and adjectives, and fixed forms of phrases and sentences which we consider as compound adverbs. To support our explanation, the morphological and syntactic properties of adverbs are discussed with typological and language universal point of view. Semantic classification is not included in this article because we think it can not be made rigorously without delimiting and analyzing possible various forms of adverbs(and adverbials) and their syntactic functions.
Thesauri have been widely used in bibliographic databases for 30 years. Recently, CD-ROMs of a variety of dictionaries with their GUIs are spreaded to current users. On the other hand, End users dose not use thesauri as for their bulky printed matter. The king of software browsing graphically and managing thesauri does not appear in PC environment.. The browsing tool for lexical database with hierarchical structure has been developed using Java.. This paper describes the functions, the components, and examples of its usage. The problems of the browser and the functions to be extended are discussed.
In this paper, we propose an error correction method using text corpora. In this method, recognition errors are corrected using phonetically similar examples in the text corpora. The reliability of the correction hypotheses are judged according to their semantic consistency and their phonetic similarity to the original input. We previously proposed an error correction method that uses a treebank [1]. However, the previous method was not flexible in its use of examples, because structural mismatches occurred between the input and examples due to recognition errors. In our new proposal, examples are treated as morpheme sequences. This enables us to use examples partially when there are no useful full-sentence-examples. We built our proposed method into a speech translation system and compared the translation quality for simple translation and translation with error correction. The rate of acceptable translation increased about 10% with our proposed method compared to simple translation.
Multiway trees (MT, henceforth) are a common and well-understood data structure for describing hierarchical linguistic information. With the availability of large treebanks, retrieval techniques for highly structured data now become essential. In this contribution, we investigate the efficient retrieval of MT structures at the cost of a complex index---the Treegram Index.We illustrate our approach with the VENONA retrieval system, which handles the BHt (Biblia Hebraica transcripta) treebank comprising 508,650 phrase structure trees with maximum degree eight and maximum height 17, containing altogether 3.3 million Old-Hebrew words.
The use of the Inside-Outside (IO) algorithm for the estimation of the probability distributions of Stochastic ContextFree Grammars is characterized by the use of all the derivations in the learning process. However, its application in real tasks for Language Modeling is restricted due to the that it needs to converge. Alternatively, several estimations algorithms which consider a certain subset of derivations in the estimation process have been proposed elsewhere. This set of derivations can be chosen according to structural criteria, or by selecting the k-best derivations. These alternatives are studied in this paper, and they are tested on the corpus of the Wall Street Journal processed in the Penn Treebank project.
This study examined the effects of performance information from the ratee's coworker both on rater's perceptions of the information and ratings of the ratee's performance. As part of a managerial role-play exercise, subjects were required to rate an employee's performance. Audiotaped interactions of the ratee and customers were manipulated to reflect either good or poor performance. In addition, the ratee's coworker provided performance information which varied on (a) who was the initiator of the report and (b) the favorability of the information. Results revealed that while performance information from the ratee's coworker did not significantly affect ratings of performance when it was consistent with the rater's direct observations, ratings were affected when the information from a coworker was inconsistent with the rater's direct observations even though the information from a coworker was perceived as less accurate and of less use. Copyright © 1999 John Wiley & Sons, Ltd.
We present a new approach to partial parsing of natural language texts that relies on machine learning methods. The approach combines corpus-based grammar induction with a very simple pattern-matching algorithm and an optional constituent verification step. The grammar induction algorithm acquires a set of rules for each level of linguistic analysis using a new technique for errordriven pruning of treebank grammars. The constituent verification step employs standard inductive learning techniques as an additional precision-enhancing device. We evaluate the approach on four partial parsing data sets and find that performance is very good (over 93% precision and recall) for applications that require or prefer fairly simple constituent bracketing. As the complexity of the partial parsing task increases, however, our approach lags the performance of competing approaches. We explain these differences in terms of the knowledge sources employed by each method and describe a number of features...
Discusses the linguistic influences on an electronic publishing infrastructure in an environment with unstable linguistic standardization from the computational point of view. Essentially, in Serbia in the last half of the century (at least) publishing is based on the following facts: two alphabetic systems are regularly in use with the possibility to mix both alphabets in the same document; the various dialects are accepted as a part of a linguistic norm; orthography is unstable ‐ presently, several linguistic attitudes that have different views of the orthographic norm are under discussion; and, in Serbia, many minority languages are in use, which makes it difficult to provide efficient contact between different communities through electronic publishing. In this context, a systematic solution that responds to this complex situation has not been developed in the frame of traditional Serbian linguistics and lexicography in a way that enables the adequate incorporation of the new publishing technologies. Owing to these constraints, the direct application of electronic publishing tools frequently causes the degradation of the linguistic message. In such an environment, the promotion of electronic publishing therefore needs specific solutions. The paper discusses the general frame based on the specifically encoded system of electronic dictionaries that makes electronic texts independent of some of the mentioned constraints. The objective of such a frame is to enable the linguistic normalization of texts at the level of their internal representation, and to establish bridges for communicating with other language societies. Some aspects of electronic text representation that ensures its correct interpretation in different graphical systems and in different dialects are described. This also allows text indexing and retrieval using the same techniques that are available for languages not burdened with these problems.
En esta tesis se estudian las Gramaticas Incontextuales Probabilisticas y su aplicacion en problemas de Modelizacion del Lenguaje, Dos son los grandes problemas que se va a considerar en este tipo de modelos: el aprendizaje de las funciones de probabilidad asociadas a las reglas, y su integracion como modelo de interpretacion en tareas complejas de Modelizacion del Lenguaje. En primero de los problemas que se estudia es la estimacion de las funciones de probabilidad asociadas a las reglas.Se presentan y estudian dos de los algoritmos clasicos de estimacion de las GIP, el algoritmo Inside-Outside y el algoritmo basado en las cuentas de Viterbi. Despues se proponen nuevos algoritmos de estimacion en los cuales se utiliza un subconjunto especifico de derivaciones e cada cadena. Finalmente,los algoritmos propuestos se aplican al conjunto de datos del Penn Treebank para ilustrar su comportamiento en la practica. Por utlimo se aborda el problema de la interpretacion e integracion de las Gramaticas Incontextuales Probabilisticas en problemas de Modelizacion del Lenguaje. A continuacion se hace una propuesta de integracion que combina modelos de $n$-gramas a nivel de palabras con una Gramatica Incontextual Probabilistica a nivel de categorias lexicas.
We tested new analytic procedures for combining an observer's image-ratings of lesion-likelihood with localization reports that are incomplete (unavailable on images rated as 'normal') and/or imprecise (possibly scored as 'correct' by chance), and for fitting a constrained ROC formulation to the rating data alone. Eight radiologist readers in a previous study had rated the likelihood of nodular lesions on each of 250 chest-film cases (39 with subtle nodules, 36 with 'typical' nodules and 175 normal cases) that were presented in two display modes (original films or on video workstation). Ratings in the four positive categories (2 to 5) were accompanied by reports that grossly localized the suspected nodules into one of 7 film- regions (upper, middle or lower portions of left or right lung field, or retrocardiac), but there was no localization for the cases rated as 'normal' (category 1). In each of 29 sets of data, we estimated the area below the ROC curve (A<SUB>z</SUB>) and its standard error using three different fits: (1) the usual ROC formulation, (2) the constrained ROC formulation and (3) the new procedure that included incomplete and imprecise localization data (I&I). Estimates of A<SUB>z</SUB> from the usual and constrained ROC fits were quite similar unless the standard ROC exhibited an upward 'hook,' but standard errors of A<SUB>z</SUB> were always the same or smaller for the constrained ROC fit. The I&I fit that included localization data often estimated A<SUB>z</SUB> to be either larger or smaller than the usual or constrained ROC fits that considered only the rating data, but its A<SUB>z</SUB> had substantially smaller standard errors in 28 of the 29 sets of observer data.
This paper challenges traditional notions of plagiarism as an act of academic deviance and suggests that, particularly for LBOTE students.'plagiarism'is often a text-based practice that reflects different cultural and linguistic norms. The author suggests that western institutions' ambivalent and inconsistent approach to the practice further clouds the issue for students and teachers
中文句結構樹資料庫(Sinica Treebank)建構的主要目的是提供中文自然語言處理研究一個具有標記語料庫的研究素材,我們可以從這個中文句結構樹資料庫中抽取語法知識,也藉由語法知識的抽取與瞭解使我們的剖析系統功能更趨完善。本文介紹中文句結構樹資料庫構建方法和步驟,從五百萬詞的中央研究院平衡語料庫(Sinica Corpus),抽取句子,以訊息為本格位語法(Information-based Case Grammar, ICG)的表達模式為基本架構,經由電腦自動剖析成結構樹,可以盡量維持結構標記的一致性,最後並加以人工修正、檢驗,以維持標記的正確性。對於歧義的句法結構形式及詞類標記,我們也提出處理的原則。
本稿では, ツリーバンクを用いて入力文と類似した文の構文木から入力文に対する構文木を類推する手法を提案する. この手法は用例に基づく解析手法の1つであるが, 統計情報や意味的類似性ではなく, 複数のツリーバンク内データの問で定義される特定の類似関係に基づいて構文解析を行う. 特にここではツリーバンク内の知識表現形式をそのまま使って構文解析を行うため, 比較的容易に他の解析手法との融合を考えることができる. またこの手法は辞書などを用いず, データ間の類似性のみに基づいて解析を行うため, 未知語などを含む入力に対しても頑健に働く. ここでは特に基本原理として働く類似関係の有効性を評価するためにPenn Treebankを用いて評価実験を行った. その結果, 単語の表層情報と品詞情報を用いることで解析可能な文の約70%が一意に正しく解析でき, また誤ったものについても比較的正解に似た構文木を出力することができた.
In order to preserve the basic unity of a language like Spanish, which is spoken in 20 countries, all the speakers must adopt a respectful and careful attitude when they use it orally. In the area of phonetics, educated Spanish speakers, even those who study the language, express themselves in a somewhat careless way and what is most surprising is that such negligence is accepted in Spain by the educated linguistic norm.
We examined male and female abservers' reactions to the use of touch by a nurse towards a patient in a hospital situation. The results suggest that an overrall pattern for observers to react more favorably to the nurse touch compared to no touch in interacting with patients may be judged in part by the attitudes of males and females about the use of touch. The generally favorable reaction to the use of touch by a nurse is consistent with Lewis et al's (1995) study. However, in contrast to their study, there were no differences in the nurse's supportiveness, competence, affect ratings and nurse's confidence in the low touch and high touch conditions.
Hebrew Studies 40 (1999) 269 Reviews grams that serve as excellent visual aids. The work is written in a highly technical language that assumes familiarity with the terminology. shorthand, and conceptual framework of Generative Grammar. and this limits its accessibility and appeal to a wide readership. The inclusion of a glossary andlor a brief overview of this method of linguistic study would help address this problem. The book's title is somewhat inaccurate since it does not offer a truly comparative study of Hebrew and Arabic syntax. It is principally an analysis of Hebrew which makes reference to other languages to explicate and illustrate Shlonsky's ideas on Hebrew. While Arabic is cited more frequently than any other language besides Hebrew. there are lengthy sections of the book in which it never figures in the discussion. For example. there is not a single reference to Arabic in the thirty-page chapter treating subject-verb inversion. A further problem concerns the inconsistent manner in which Shlonsky appeals to the Arabic evidence. Throughout the book, he argues his case by referring to several different forms of the language. including Standard Arabic and the dialects from Palestine. Southern Palestine. Morocco, and Egypt. Each of these is a unique linguistic system that is distinct from the others but Shlonsky does not pay sufficient attention to the differences among them. This method hinders the purported comparative focus of his work since the reader is not sure which Arabic is meant to be the primary point of comparison. Consistent reliance upon one form of the language would have been a more beneficial approach to adopt. Such relatively minor problems do not significantly detract from the many strengths of this volume and the important contribution it makes. Shlonsky's book is required reading for anyone interested in serious study of Hebrew and will be a major work in the field. John Kaltner Rhodes College Memphis, TN 38112 kaltner@rhodes.edu THE SEMANTICS OF ASPECT AND MODALITY: EVIDENCE FROM ENGLISH AND BIBLICAL HEBREW. By Galia Hatav. Studies in Language Companion Series 34. Pp. x + 224. Philadelphia. PA: John Benjamins, 1997. Cloth, $85.00. Originally a dissertation at Tel-Aviv University. this work "aims to provide a general (semantic) theory for temporality...but it also systemati- Hebrew Studies 40 (l999) 270 Reviews cally examines the verb system in Biblical-Hebrew...which lacks tenses, as will be demonstrated, and thus enables us to see the nature of aspect and modality more clearly" (p. I). The author's theoretical assumption is "that TAM, i.e., the Tense-Aspect-Modal system in language, should be defmed within truth conditional semantics, in terms of temporality, rather than within a pragmatic approach which deals with it in terms of perspective, attitude, and the like" (although pragmatics is not ignored altogether, p. 195). Moreover, she seeks to analyze the data within the framework of a threefold distinction. Here she builds on the work of Hans Reichenbach, who argued that the contrast between the time of speech (S-time) and the time in which the event actually took place (E-time) is not sufficient to account for verbal uses. A third category is needed, namely, the time of reference (R-time), a somewhat fuzzy concept that Hatav defmes as a time unit that contains (or is contained in, or is ordered with) the E-time (pp. 3-5). In the introductory chapter the author, after explaining these and other assumptions, surveys previous attempts to account for the verbal system of biblical Hebrew; unfortunately, she seems unaware of Bruce K. Waltke and M. O'Connor, Biblical Hebrew Syntax, which gives considerable attention to verbal aspect She further informs us that she examined sixty-two chapters taken from the Pentateuch and the Former Prophets (excluding poetic material, since "poetry often violates otherwise valid linguistic norms," p. 24) and gives us some details regarding her method. Following the lead of discourse-analysis scholars, such as R. Longacre, she argues that "the biblical Hebrew system organizes the text into sequential and non-sequential material," but that in addition to sequentiality, three other temporal parameters are needed. These four parameters are individually considered in the following chapters. Chapter 2, accordingly, deals...
In order to preserve the basic unity of a language like Spanish, which is spoken in 20 countries, all the speakers must adopt a respectful and careful attitude when they use it orally. In the area of phonetics, educated Spanish speakers, even those who study the language, express themselves in a somewhat careless way and what is most surprising is that such negligence is accepted in Spain by the educated linguistic norm.
1019 Aminophylline-based creams are marketed as fat reducing agents for the thighs. The purposes of this study were to determine if: 1) a thigh reducing cream would influence body image, 2) product price would influence perception of effectiveness and, 3) the product would decrease thigh size. Serving as their own control, 11 women with thigh cellulite were randomly assigned to a double-blinded, counterbalanced cream treatment with 2% aminophylline (A) on one leg and a placebo (P) on the other (age: 26±7 yrs.; BMI: 23±2 kg/m2; Body Fat (BF): 24±4%; VO2: 39±4 ml/kg/min). They were also randomly assigned into a fictitious expensive (E) or inexpensive (I) cream treatment group. In the lab, subjects massaged 4.5 g of either A or P cream for 1 min into each thigh 5 d/wk for 6 wks. Pre/post testing was done 1 wk prior and immediately after the 6 wk intervention. Dependent measures included thigh girth and skinfolds (distal, mid, proximal to patella) and psychological indices (mood, body image, rating of effectiveness). Results: No significant differences (NSD) were found at baseline between the E vs. I groups for VO2, %BF, or BMI (Independent T's, p>.20). NSD were found for the right vs. left thigh measures at baseline, nor for pre-post measures of VO2, %BF, or BMI (Paired T's, p>.22). NSD were found for skinfold and girth measures between A vs. P-treated thighs (Paired T's, p>.10). Subjects in both the E and I groups reported improved thigh image (Friedman 2-way ANOVA, p=.046), but neither the E nor I group were influenced by product cost (Mann-Whitney U-Wilcoxon Rank Sum Test, p>.17). Conclusions: The 2% A cream was not effective in reducing thigh size, however, the subjects felt more positive about their thighs and overall body build. (Partially supported by FAU Foundation)
A new ambiguity representation scheme SPR(structure preference relation) is proposed in this paper, which consists of useful quantitative distribution information for ambiguous structures. Two automatic acquisition algorithms, (1) acquired from treebank, (2) acquired from raw texts, are introduced, and some experimental results which prove the availability of the algorithms are also given. At last, some SPR applications in linguistics and natural language processing are introduced and some future research directions are proposed in this paper.
767 Electroencephalographic (EEG) frontal asymmetry has been proposed as both a physiologic index and a diathesis or disposition for emotional responding (Davidson, 1992; 1994). This study used a counter-balanced repeated measures design to examine the effect of 30 minutes of cycling exercise at ∼50%V̇O2peak, compared to 30 minutes of rest, on changes in emotional response ∼30 minutes after condition to standardized negative and positive images presented by slide projection (International Affective Picture System, Lang et al., 1988a). Emotional response was measured by EEG and self-ratings of valence and arousal (Self-Assessment Manakin, Lang et al., 1988b) in 13 females and 19 males (23±3 and 25±4y) having moderate levels of cardiorespiratory fitness (V̇O2peak = 41±9 and 53±11 ml·kg−1·min−1, respectively). Consistent with Davidson's diathesis model, response frontal α (lognμV2/Hz) right-minus-left asymmetry was predicted by resting α asymmetry, independent of neutral response (negative: t(29)=1.72, p=.09; positive:t(29)=2.35, p<.03). Valence ratings were weakly related to resting α asymmetry (p<.03). The cycling condition did not alter emotional response as indicated by [1] α asymmetry (negative: F(1,28)=0.04, p=.84, η2=.002; positive: F(1,28)=0.48, p=.49, η2=.017), [2] valence (negative: F(1,28)=0.23, p=.64, η2=.008; positive: F(1,28)=0.88, p=.36, η2=.030), or [3] arousal (negative: F (1,28)=0.15, p=.71, η2=.005; positive: F(1,28)=0.18, p=.67, η2=.007). The β1 and β2 EEG spectra were similarly unaffected by condition. Also, cycling exercise failed to alter positive or negative affect (PANAS, Watson et al., 1988), but it decreased state anxiety (STAI-Y1, 10-item)(F(1,28)=6.56, p=.02, η2=.190). Results did not differ according to gender. In the cycling condition, state anxiety change was not influenced by trait anxiety (STAI-Y2) or pre-condition resting α power asymmetry (F(2,29)=1.21,p=.31, R2adj=.014). These results indicate that moderate intensity cycling exercise lasting 30 minutes does not alter emotional response as measured, despite reducing state anxiety.
Behaviour studies were conducted on 12 adult rams of Garole breeds which were transported from hot humid climate of Sunderban area of West Bengal to semi-arid climate of Avikanagar in Rajasthan. The observations were recorded on their agnostic, investigatory, precopulatory, mounting and ejaculatory behaviour with a view to assess dominance, libido and serving capacity of these rams under semi-arid environment. The dominance rating of these rams were recorded on the basis of incidences of chasing, bunting and fights. The dominant rams had higher frequency and higher number of investigations of ewes. The incidence of anogential sniffing and frequency of bouts of leg kicking were also higher in dominant rams. The latency period, the number of mounts between ejaculations and time taken for first, second and third ejaculations were consistently lower in dominant rams.
Absract Hanne Ruus's thesis on Central Parts of The Danish Lexical Norm is a very beautiful book. The author and the publisher have bestowed great care on these two volumes and should be proud of the result. It is a pleasure to take them in hand and consult them: tatsteful, nice cover and binding, exquisite paper quality, excellent typography and layout.
Over the years, many proposals have been made to incorporate assorted types of feature in language models. However, discrepancies between training sets, evaluation criteria, algorithms, and hardware environments make it difficult to compare the models objectively. In this paper, we take an information theoretic approach to select feature types in a systematic manner. We describe a quantitative analysis of the information gain and the information redundancy for various combinations of feature types inspired by both dependency structure and bigram structure, using a Chinese treebank and taking word prediction as the object. The experiments yield several conclusions on the predictive value of several feature types and feature types combinations for word prediction, which are expected to provide guidelines for feature type selection in language modeling.
This article examines the reliance of U.S. campuses on international teaching assistants (ITA) for staffing undergraduate course and the strategies that may affect ratings of their speaking competence. This increasing reliance has led to student complaints about incomprehensibility of ITA. This problem has been examined by looking through the eyes of the students, administrators and taxpayers. Therefore, the responsibility had been placed on the ITA, whose burden it was to learn the language and culture more fully. The goal of having ITA learn the language and culture better was eclipsing another important issue that needs consideration, the issue of teaching assistant's feelings of loss of control over the students' perception about themselves.
Research has shown that seeing another person (i.e., a model) perform at a certain level can influence the goal choice and performance of an observer. This study extended these findings by examining the model's effects on two potential explanatory mediators: expectations of reaching different performance levels and valence at those levels. The moderating effects of observer task experience and self‐esteem were also examined. In a repeated measures design, results showed that model performance influenced observers' goals, task performance, expectancies, and valence ratings. Path analyses indicated that expectancies mediated observational effects on goal choice. Results also indicated significant Model x Trial interactions, with a diminishing effect of model influence as personal experience developed. Results are discussed in terms of mediating effects of expectancies and valences on model influences on goal choice as well as the different weights given to social and personal information.
This paper explores the automatic construction of a multilingual Lexical Knowledge Base from preexisting lexical resources. We present a new approach for linking already existing lexical/semantic hierarchies. The Relaxation labeling algorithm is used to select --among all the candidate connections proposed by a bilingual dictionary-- the right connection for each node in the taxonomy. We also propose several ways in which this technique could be applied to enrich and improve existing lexical databases.