Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
16504 papers
ABSTRACT This paper seeks to explore the nature of certain types of interaction between the source and the target language in the process of translation basing on the notion of language interference in the sense of any violation of the target language form or norm under the influence of the source language form or norm. It is suggested that apart from 'overt' manifestations of interference, that can be easily traced back to the source language, there also exists another type of interference, or 'cryptic' interference. Its mechanism consists in switching between the source and the target language, which affects the mental processing and results in producing instances of interference which can be traced back to the source language, however, not to the source language phrases, formulations, and lexical items found in the actual source language text; these phenomena are due to the process of re-analysis (or reformulation) of the source language message occurring before the actual translation is performed. 1. Introduction The aim of this paper is to present the findings of an empirical study of the process of sight translation, based on an experiment during which eight translators with varying professional experience were asked to perform the same translational task. The goal of the experiment was to inspect the issues connected with language interference during interpreting. (1) Naturally, language interference can be a significantly variable factor and it is reasonable to assume that both its scope and amount change relative to interpreting conditions, e.g. whether a translation task is performed from or into one's native language. The case in point here was to provide some insights into the mechanism of Foreign Language (FL) interference and the influence it exerts over the Native Language (NL) interpreting performance, as well as into the interpreters' shortcomings in their mother tongue. After inspection of the results of the experiment it turned out that a number of translation errors (2) could not be ascribed to language interference as it is understood in foreign language teaching. (3) The subjects had certain problems also with the use of their native language, a phenomenon of large significance, since all of the experiment participants had Polish as their mother tongue. Most importantly, however, the experiment revealed that regardless of their professional experience, all translators produced instances of interference of a particular character, given the working name cryptic interference. This paper is devoted to the description of its mechanism and the implications it may have for the mental aspects of the process of translation. 2. The experiment The subjects of the experiment were two groups of four. The first group featured active interpreters with professional experience varying from 18 to 3 years in the trade. The second group featured individuals with little (or no) professional experience in interpreting, but with two years' training in interpreting. It was hoped that inclusion of interpreters with varying professional experience would broaden the research spectrum and thus help render more comprehensive results. The experiment was made up of two sight-translation exercises: two attempts at interpreting one text, the first one after a careful perusal of the Second Language (SL) text, the second one after studying a text used as a prompt for the second performance--a model translation of the original SL text. For this purpose an experienced professional translator (c. 20 years in the trade) had been asked to produce a model translation. The text used in the experiment was carefully selected as it had to meet a number of requirements. It was to contain a selection of idiomatic phrases, stylistically specific for English and thus assumed to be difficult to render in a foreign language. The text selected was an authentic text, a news report from The Guardian Weekly of 25 April 1999. …
This empirically based study aims at establishing the frequency, distribution, tenacity and the nature of errors produced by more than 700 advanced Swedish learners of French, at different university levels. The errors are classified in grammatical and lexical categories. Four grammatical categories (articles, nouns, pronouns, verbs) have been selected for a thorough examination, and examples of different types of errors are presented and commented upon. The error analysis follows, principally, the steps suggested by Corder (1974) and implies the discussion of such concepts as error, norm and usage.The verbs present the most frequent and tenacious errors constituting at each level about a third of all errors. The proportions of the different types of errors change according to task and level. The errors concerning tenses predominate in the translation tests, those concerning conjugations in the diagnostic test. As for pronouns, they cause 40 % of all errors in the diagnostic test (fill-in/transformation tests), 7 %-8 % in the translation tests. This reduction is partly due to the fact that many pronouns are necessarily regarded as part of the construction of a verb and classed as such. Put together, the errors concerning gender, orthography and vocabulary constitute 40 % at the advanced levels. The study includes a correlational analysis aiming at confirming or confuting some hypotheses regarding the interrelation, on the one hand between the results of the different parts of the two tests given during the first term of the university studies (vocabulary, grammar, metalingual knowledge), on the other hand beween the results and some external factors such as marks in French (secondary school), stay in French-speaking country, length of previous studies of French, etc. Among other things, it is shown that the students’ results in grammar are systematically better than those in vocabulary and that six years of previous studies of French do not generally lead to better results than three years’ studies.
The author's research shows that nicknames are one of the oldest antroponemic categories as well as one of the most affective ones because nicknames usually keep both onomastic and lexical meaning that have motivated them. Nicknames are, more than other names, a product of speech. They are the act of speech and are not subject to the language norm. Nicknames bear message and mocking in their contents depending on the person they refer to. Due to those speech characteristics, nicknames fit into the topic of this journal that has speech in its title, and follow the huge, greatly admired and prominent opus of the distinguished colleague full professor Ivo Skaric to whom I dedicate this paper on the occasion of his birthday and the anniversary of his scientific work.
C-rater is an automated scoringengine that has been developed to scoreresponses to content-based short answerquestions. It is not simply a stringmatching program – instead it uses predicateargument structure, pronominal reference,morphological analysis and synonyms to assignfull or partial credit to a short answerquestion. C-rater has been used in two studies:National Assessment for Educational Progress(NAEP) and a statewide assessment in Indiana.In both studies, c-rater agreed with humangraders about 84% of the time.
In this text we present``profile-based linguistic uniformity'', a methoddesigned to compare language varieties on thebasis of a wide range of potentiallyheterogeneous linguistic variables. In manyrespects a parallel can be drawn with currentmethods in dialectometry (for an overview, see,Nerbonne and Heeringa, 2001; Heeringa, Nerbonneand Kleiweg, 2002): in both casesdissimilarities between varieties on the basisof individual variables are summarized inglobal dissimilarities, and a series oflanguage varieties are subsequently clusteredor charted using multivariate techniques suchas cluster analysis or multidimensionalscaling. This global similarity between themethods makes it possible to compare them andto investigate the implications of notabledifferences. In this text we specifically focuson, and defend one characteristic of ourmethodology, its profile-based nature.
There are two important strategies incomputer-assisted reading and analysis of text(CARAT). The first relates to theclassification process, and the second pertainsto the categorisation process. These twooften-interrelated operations have beenregularly recognised as essential components oftext analysis. However, the two operations arehighly time-consuming. A possible solution tothis problem calls upon more inductive orbottom-up strategies that are numerical andstatistical in nature. In our own research, wehave been exploring a few of these techniquesand their combination. We now know, through ourown past research and others' work, that theclassification methods allow a good empiricalthematic exploration of a corpus. Morespecifically, in this paper we shallconcentrate on the problem of assisting theautomatic categorisation of small segments of aphilosophical text into a set of thematiccategories.
222 Reviews Bullock: 'The advantage to this kind of constraint based approach over rule based approaches is that it obviates the need forderivations in phonology' (p. 55). Such opposing assertions prove that, happily, aftermore than a decade, something is starting to move in phonological theory. But, if they keep within the domain of untestable formal constructs, they are little more than rhetorical exercises of fittingdata into theories?instead of searching for a theory to fitthe data. The clearest conclusion from this book may be, perhaps, that phonological theory has much to gain from contact with Italian dialectology. It is true that most dialectologists have tried to work in isolation from modern (generativist) theories. This has been regrettable for dialectologists themselves, who may have missed important generalizations, but also for theoreticians, who have tended to produce increasingly abstract artefacts. But, as this volume proves, even if Italian generative dialectologists may be considered small in number, that is not coextensive with saying that Italian theory-oriented dialectologists are not many. Let us hope that this book might be a step on the road for Romance dialectology to regain the leading role in theoretical innovation that it had a century ago, at the time of Gillieron and Gauchat. University of Salamanca Carmen Pensado Two Spanish Songbooks: The 'Cancionero Capitular de la Colombina' (SV2) and the 'Cancionero de Egerton' (LB3). Ed. by Dorothy Sherman Severin; editorial assistant Fiona Maguire. (Hispanic Studies TRAC, 11) Liverpool and Seville: Liverpool University Press and Institucion Colombina. 2000. 438 pp.?47.95 (pbk?22.95). ISBN 0-85323-650-x (pbk 0-85323-109-5). This edition oftwo fifteenth-centurycancioneros provides wider access to manuscripts until now available only in the partial edition of Brian Dutton (El cancionero del siglo XV (c. 1360-1520), ed. by Brian Dutton, musical cancioneros ed. by Jineen Krogstad, 7 vols (Salamanca: Biblioteca Espanola del Siglo XV and Universidad de Salamanca, 1990, 1991), where incomplete editions of SV2 and LB3 appear in vols iv, 301-14, and 1,359-72, respectively. The Introduction (pp. 1-34) consists of a description and brief analysis of the contents of both manuscripts, followed by an index which supplies the Dutton ID number foreach composition, details of its metrical structure (if appropriate), author, title (where relevant), and first line or first stanza. Next comes an account of the genealogical relationship of the two cancioneros derived from earlier criticism. In the Norms of Transcription the editor expresses the aim of providing 'a diplomatic, readable version ofthe two texts and not a critical edition with extensive critical apparatus' (p. 31). This section concludes with a bibliography. The nucleus of the edition consists of the texts of SV2 (pp. 35-276) and LB3 (pp. 277-430). Both texts are provided with abundant palaeographic notes at the end of each section. An Author Index (pp. 431-34) and an index of first lines or first stanzas completes the volume. In the form in which they are presented, the texts represent an adequate fulfilment ofthe objectives articulated by the editor: 'to make a corpus of hitherto unedited texts of cancioneros available to students and scholars' (p. 31). The question arises, however, as to whether an edition directed not only at scholars but also at students might not also profitably include a minimum number of footnotes devoted to the lexical, histor? ical, and literary background, since it is improbable that all students will be familiar with terms such as adarve, blanchetes, or xorginos, or able to place either Macias or Juan de Merlo in a precise context. The editorial method seems more directed at the experienced medievalist than the beginner. MLR, 98.1, 2003 223 Greater attention to these considerations would have allowed the editor to avoid some inexactitudes, e.g. in the faulty placing ofthe reference [ID0091 P0050]. Dut? ton' s reference is to Santillana's prose prologue to his Proverbios, and not to Pero Diaz de Toledo's preface to his commentary on Santillana's poem. Severin appears to confuse the two, however, misplacing the Dutton reference, which should appear on page 47 of the edition before the paragraph which heads Santillana's work '[S][ere]nisymo &bienaventuradoprincipe[...]', rather than on page...
Gender difference in language is a universal phenomenon which not only embodies the language user's cultural psychology and social values but also reflects the social norms and ethnography. This article discusses and contrasts, from the sociolinguistic perspective, the gender differences as manifested in English phonology, lexical choice, syntactic structure and language communication. The discussion aims at making a scientific, accurate and objective explanation of the gender differences in language use.
This paper presents a numeric and information theoretic model for themeasuring of language change, without specifying the particular type ofchange. It is shown that this measurement is intuitively plausibleand that meaningful measurements canbe made from as few as 1000 characters. This measurement techniqueis extended to the task of determining the ``rate'' of language changebased on an examination of brief excerpts from the NationalGeographic Magazine and determining both their linguistic distancefrom one another as well as the number of years of temporal separation.A statistical analysis of these results shows, first, that language changecan be measured, and second, that the rate of languagechange has not been uniform, and that in particular, the period 1939-;1948had particularly slow change, while 1949-;1958 and 1959-;1968 hadparticularly rapid changes.
The computation of the optimal phonetic alignment andthe phonetic similarity between wordsis an important step in many applications in computational phonology,including dialectometry.After discussing several related algorithms,I present a novel approach to the problem that employsa scoring scheme for computing phonetic similarity between phonetic segmentson the basis of multivalued articulatory phonetic features.The scheme incorporates the key concept of feature salience,which is necessary to properly balance the importance of various features.The new algorithm combines several techniquesdeveloped for sequence comparison:an extended set of edit operations,local and semiglobal modes of alignment,and the capability of retrieving a set of near-optimal alignments.On a set of 82 cognate pairs,it performs better than comparable algorithms reported in the literature.
This paper considers the question of authorship attribution techniques whenfaced with a pastiche. We ask whether the techniques can distinguish the real thing from the fake, or can the author fool the computer? If the latter, is this because the pastiche is good, or because the technique is faulty? Using a number of mainly vocabulary-based techniques, Gilbert Adair's pastiche of Lewis Carroll, Alice Through the Needle's Eye, is compared with the original `Alice' books. Standard measures of lexical richness, Yule's K andOrlov's Z both distinguish Adair from Carroll, though Z also distinguishesthe two originals. A principal component analysis based on word frequenciesfinds that the main differences are not due to authorship. A discriminantanalysis based on word usage and lexical richness successfully distinguishes thepastiche from the originals. Weighted cusum tests were also unable to distinguish the two authors in a majority of cases. As a cross-validation, wemade similar comparisons with control texts: another children's story from thesame era, and other work by Carroll and Adair. The implications of thesefindings are discussed.
A course that relies on open-source software for teaching introductory computer programming and Web development to psychology graduate and advanced undergraduate students is described. The rationale, content, learning goals and outcomes of the course are described, along with the specific software used. The advantages of relying on open-source solutions rather than commercial software for implementing such a course are discussed.
In this study, we investigated whether computer-animated graphics are more effective than static graphics in teaching statistics. Four statistical concepts were presented and explained to students in class. The presentations included graphics either in static or in animated form. The concepts explained were the multiplication of two matrices, the covariance of two random variables, the method of least squares in linear regression, α error, β error, and strength of effect. A comprehension test was immediately administered following the presentation. Test results showed a significant advantage for the animated graphics on retention and understanding of the concepts presented.
Quantile maximum likelihood (QML) is an estimation technique, proposed by Heathcote, Brown, and Mewhort (2002), that provides robust and efficient estimates of distribution parameters, typically for response time data, in sample sizes as small as 40 observations. In view of the computational difficulty inherent in implementing QML, we provide open-source Fortran 90 code that calculates QML estimates for parameters of the ex-Gaussian distribution, as well as standard maximum likelihood estimates. We show that parameter estimates from QML are asymptotically unbiased and normally distributed. Our software provides asymptotically correct standard error and parameter intercorrelation estimates, as well as producing the outputs required for constructing quantile—quantile plots. The code is parallelizable and can easily be modified to estimate parameters from other distributions. Compiled binaries, as well as the source code, example analysis files, and a detailed manual, are available for free on the Internet.
A common tool for improving theperformance quality of natural languageprocessing systems is the use of contextualinformation for disambiguation. Here I describethe use of a finite state machine (FSM) todisambiguate speech acts in a machinetranslation system. The FSM has two layers thatmodel, respectively, the global and localstructures found in naturally-occurringconversations. The FSM has been modeled on acorpus of task-oriented dialogues in a travelplanning situation. In the dialogues, one ofthe interactants is a travel agent or hotelclerk, and the other a client requestinginformation or services. A discourse processorbased on the FSM was implemented in order toprocess contextual information in a machinetranslation system. Evaluation results showthat the discourse processor is able todisambiguate and improve the quality of thedialogue translation. Other applicationsinclude human-computer interaction andcomputer-assisted language learning.
Responses in personalinterviews about education and career with 415Swedish men and women (age 34) forms the basisof a speech corpus with 1.8 million words. Thevocabulary is described by means of two sets ofvariables. One is based on the number of tokensand types, word length and sectioning of therunning text. The other set divides the corpusinto grammatical categories. Both sets ofvariables are related to a number of backgroundvariables such as gender, socioeconomicbackground, education, and indicators of verbalproficiency at age 13 and 32. This possibilityto study the relationship between vocabularyand a broad set of respondent characteristicsis a unique feature of this corpus.
The MIRID CML program is a program for the estimation of the parameter values of two different componential IRT models: the Rasch—MIRID and the OPLM—MIRID (Butter, 1994; Butter, De Boeck, & Verhelst, 1998). To estimate the parameters of both models, the program uses a CML approach. The model parameters can also be estimated with a MML approach that can be implemented in PROC NLMIXED of SAS Version 8. Both the MIRID CML program and the MML SAS approach are explained and compared in a simulation study. The results showed that they did about equally well in estimating the values of the item parameters but that there were some differences in the estimation of the person parameters, as could be expected from the differential assumptions regarding the distribution of the persons. The SAS MML approach is much slower than the MIRID CML program, but it is more flexible.
We investigated the reliability and validity of a video-based method of measuring the magnitude of children’s emotion-modulated startle response when electromyographic (EMG) measurement is not feasible. Thirty-one children between the ages of 4 and 7 years were videotaped while watching short video clips designed to elicit happiness or fear. Embedded in the audio track of the video clips were acoustic startle probes. A coding system was developed to quantify from the video record the strength of the eye-blink startle response to the probes. EMG measurement of the eye blink was obtained simultaneously. Intercoder reliability for the video coding was high (Cohen’sκ = .90). The average within-subjects probe-by-probe correlation between the EMG- and video-based methods was .84. Group-level correlations between the methods were also strong, and there was some evidence of emotion modulation of the startle response with both the EMG- and the video-derived data. Although the video method cannot be used to assess the latency, probability, or duration of startle blinks, the findings indicate that it can serve as a valid proxy of EMG in the assessment of the magnitude of emotion-modulated startle in studies of children conducted outside of a laboratory setting, where traditional psychophysiological methods are not feasible.
In this article, we present the spatial logistics task (SLOT) platform for investigating multimodal communication between 2 human participants. Presented are the SLOT communication task and the software and hardware that has been developed to run SLOT experiments and record the participants’ multimodal behavior. SLOT offers a high level of flexibility in varying the context of the communication and is particularly useful in studies of the relationship between pen gestures and speech. We illustrate the use of the SLOT platform by discussing the results of some early experiments. The first is an experiment on negotiation with a one-way mirror between the participants, and the second is an exploratory study of automatic recognition of spontaneous pen gestures. The results of these studies demonstrate the usefulness of the SLOT platform for conducting multimodal communication research in both human-human and human-computer interactions.
Sociallexicological description of non-standard vocabulary of English and Russian military sublanguages includes notions of sociallinguistic norm, existential form of language, diglossia, national variant, literary standard, popular language, sublanguage, sociolect, lexical systems of sublanguage, military sublanguage and military sociolect, invective. Classified foundation of words stock of language is social-communicative stratification of components of literary standard and lexical popular language.
Machine translation engines draw on various types of databases. This paper is concerned with Arabic as a source or target language, and focuses on lexical databases. The non-concatenative nature of Arabic morphology, the complex structure of Arabic word-forms, and the general use of vowel-free writing present a real challenge to NLP developers. We show here how and why a stem-grounded lexical database, the items of which are associated with grammar-lexis specifications – as opposed to a root-&-pattern database –, is motivated both linguistically and with regards to efficiency, economy and modularity. Arguments in favour of databases relying on stems associated with grammar-lexis specifications (such as DIINAR.1 or the Arabic dB under development at SYSTRAN), rather than on roots and patterns, are the following: (a) The latter include huge numbers of rule-generated word-forms, which do not actually appear in the language. (b) Rule-generated lemmas – as opposed to existing ones – are widely under-specified with regards to grammar-lexis relations. (c) In a Semitic language such as Arabic, the mapping of grammar-lexis specifications that need to be associated with every lexical entry of the database is decisive. (d) These specifications can only be included in a stem-based dB. Points (a) to (d) are crucial and in the context of machine translation involving Arabic.
One morning each of us received a phone call from Ed Hovy. Are you sitting down? he asked. He told us that as a way to combat conference overload, and to promote interaction among communities, a joint conference had been proposed to combine HLT and NAACL. A diverse oversight committee had been formed, and according to Ed, this committee had been able to agree on two people -- and only two people -- as program co-chairs, because together we represented all of the vested interests. Marti was meant to represent the standards and tastes of the NAACL and the SIGIR crowds, and Mari the speech community, and both have been working on research contracts with HLT funders. Ed told us that if either of us said no, the entire enterprise would come crashing down. There are few better ways to convince busy people to become program co-chairs. Throughout the process, Ed provided the vision for and the drive behind this conference. We salute him for making this idea a reality, and for his enthusiastic and energetic phone calls that kept everything going. This is an exciting time for research in human language technologies. After years of relative calm, the field seems suddenly to be moving by leaps and bounds. Evidence of this can be found in our conference panel on Preparing for a Surprise Language (and as embodied in the short paper Desperately Seeking Cebuano). This panel will discuss the experiences of several groups of researchers, who at the behest of DARPA, acquired and developed language resources for an entirely new language within a span of only 10 days. This experiment took place in March of 2003, and the language in question was Cebuano, a language spoken in the Philippines. Participants successfully collected a large body of lexical and textual resources and developed a range of tools, including stemmers and POS taggers. (In June, DARPA will announce a new surprise language.) The existence of a variety of language resources, combined with advances in statistical analysis and modeling techniques, is resulting in fast-paced improvements in the field. parsers can now produce syntax trees for long sentences with high accuracy and great speed. Advances are starting to be made in automated semantic analysis. Great strides are being made in the sophistication and coverage of question answering systems. Speech recognition systems have achieved suficiently high accuracy that it is now possible to do retrieval, information extraction and topic tracking on spoken documents. Large and growing collections of text and speech corpora -- and the promise of much more from the web -- have enabled many of these advances. New developments in weakly supervised and unsupervised learning algorithms are critical for taking advantage of many new data sources, and hence this was chosen as a special theme of the conference. Lexical resources such as FrameNet, WordNet, PropBank, MeSH, and the Penn TreeBank also play prominent roles in HLT advances. As a field, human language technologies research should use, as motivation and guide, an understanding of the linguistic and cognitive bases of language. The invited talk by Dr. Elissa Newport, entitled Statistical language learning: Mechanisms for language acquisition in human learners, should help enlighten the community by informing us about the latest in psycholinguistic research. We received 162 submissions for full papers, of which 37 were accepted, resulting in a highly competitive acceptance rate of 22%. For the short (late-breaking) papers track, we received 80 submissions, of which 41 were accepted (2 later withdrawn). Some of these will be presented as short talks, and others as posters. Seventeen demonstrations will be shown. We were fortunate to be able to accept 15 papers that addressed the conference theme of unsupervised and weakly supervised methods. We also encouraged papers that described techniques that cross over or combine NLP, speech and/or IR, and several of the papers demonstrate this kind of crossover. The full paper reviewing was done using a two-tier system. First, two first-tier reviewers read every paper. Then a third reviewer, known as the meta-reviewer, wrote their own review. Finally, the meta-reviewer summarized these reviews and introduced additional comments. In some cases, the meta-reviewer instigated discussion among the first-tier reviewers to work out controversial issues. The meta-reviewers also attended the program committee meeting in which all the papers were discussed and acceptances were decided. For the short papers, each short paper received at least two reviews. Those papers whose reviewers disagreed, or which received middling scores, were subsequently reviewed by a member of the program committee and the program co-chairs. Paper submission and reviewing was done online using Marti's conference reviewing software (Conga), which she updated for this conference. Marti also maintained the conference website.
The effectiveness of a domain-specific latent semantic analysis (LSA) in assessing reading strategies was examined. Students were given self-explanation reading training (SERT) and asked to think aloud after each sentence in a science text. Novice and expert human raters and two LSA spaces (general reading, science) rated the similarity of each think-aloud protocol to benchmarks representing three different reading strategies (minimal, local, and global). The science LSA space correlated highly with human judgments, and more highly than did the general reading space. Also, cosines from the science LSA spaces can distinguish between different levels of semantic similarity, but may have trouble in distinguishing local processing protocols. Thus, a domain-specific LSA space is advantageous regardless of the size of the space. The results are discussedin the context of applying the science LSA to a computer-based version of SERT that gives online feedback based on LSA cosines.
1006 Reviews 'national' or Parisian counterpart, and to give a clear idea of its distinctive place in the French media landscape. Martin manages to give this overview in a very readable manner and without being superficial. He acknowledges and draws on the excellent work that has been done on individual titles, periods, and geographical areas. A particularly welcome aspect is the significant space devoted to considering newspapers as eco? nomic and social entities, highlighting not just the Citizen Hersants and the starjour? nalists but also the networks of correspondents in the smallest ofvillages, the typographers, the delivery drivers, the sellers, and the readers. Especially fascinating fromthe perspective of social history is the analysis of the evolution, content, and role of the 'avis de deces' rubric: starting as simple quasi-administrative announcements, often appearing after the funeral, these came to be used as a substitute for the individual 'faire-part', then as a signifierof social status. They were also a major source of income fornewspapers. Martin is sensitive throughout to the impact of new technologies, up to and including the Internet, and notes that the regional press has often pioneered their use in France. The volume is impressively useable: there is an accurate general index and a separate index of newspaper titles, running to over seven pages; a useful chronology; a detailed table of contents; and an annotated summary bibliography to complement the abundant and detailed notes. These tools will help a range of readers make the most of a volume that achieves its purpose and invites furtherstudy. University of Leeds Paul Rowe La Neologie en francais contemporain: examen du concept et analyse de productions neologiquesrecentes. By Jean-Francois Sablayrolles. Paris: Champion. 2000. 588 pp.?86.90. ISBN 2-7453-0275-2. This is a scholarly and thought-provoking contribution to a field which has in the past suffered from either too narrow an academic approach, or from being subject to merely anecdotal treatment in amusing collections of neologisms. Jean-Francois Sablayrolles is ambitious, and largely successful, in his attempt to link broad theoret? ical discussion to a significant body of data. After a concise history of the notion of 'neologism' in Greek, Latin, and French, he reviews the differentapproaches to the subject by French linguists and then summarizes how twentieth-century theoretical models, from the structuralists to generativists and the most recent work of Melcu'k, have dealt with the processes of lexical creativity. He notes that, generally speaking, they have been assigned a very secondary and marginal role. In the second part of the book Sablayrolles proposes his own definitions of neologisme and neologie, and examines the types of unit and process that these involve. Perennial issues such as the role of dictionaries, upon which linguists have to rely, albeit often grudgingly, for their data, and the problem of differentiating between polysemy and homonymy, are given a fresh airing. More original are the brief dis? cussion of links between politico-cultural ideology and attitudes to neologisms, and speculation on the possibility of calculating the lifespan of a neologism. In the third part of the book the author analyses and compares the data that he has gathered from his six corpora, and ends with a discussion of the differentfunctions of neologisms. These range from their attention-catching use in newspaper headlines to their role in political polemics and their largely ludic function in the work of the writers R. Jorif and Ph. Meyer. Somewhat problematic is his inclusion of a corpus scolaire, drawn from the written work of secondary-school students. Many of these examples are non-standard verb forms such as ils croivent and il a acqueri. One can argue that these are not lexical, and possibly not new. Unlike the rest of his data, these forms are probably, as he concedes, neither 'voulus' nor 'conscients' (p. 317). Surely they MLRy 98.4, 2003 1007 are either part of a system which just happens to be differentfrom the norm or an attempt to conjugate a lexical item which is simply alien to the system? Theoretically at least, Sablayrolles appears to give the status of neologism to all new forms, what? ever their source or motivation. (Perhaps intentional, conscious creation...
Cognitive behavioral therapy (CBT) for hoarding disorder (HD) has resulted in statistically significant improvements in hoarding symptoms, but gains have been modest and most participants continue to have clinically significant symptoms at post treatment. Contingency management, an empirically-supported intervention for substance use, may be effective in overcoming barriers to effective treatment of HD, such as fluctuating motivation and insight. The objective of the current open trial was to examine the potential effectiveness of contingency management for HD in the context of a cognitive-behavioral group therapy. Twenty-two patients completing 16-week CBT groups for HD were administered monthly contingency payments based on independent evaluator-rated reductions in overall in-home clutter. Mixed effects models suggested significant reductions in hoarding symptoms as measured by the Saving Inventory-Revised (SI-R; Frost, Steketee, & Grisham, 2004) and the Clutter Image Rating Scale (CIR; Frost, Steketee, Tolin, & Renaud, 2008), with SI-R reductions resulting in a large effect size (Cohen's d =2.59) that surpassed those obtained previously in trials of CBT for HD. Mean total earning per patient was $139, and ranged from $0 to $270. These preliminary results suggest that contingency management shows promise as a cost-efficient adjunctive intervention to boost gains in CBT for HD.
This study examines expressional styles and plans for the education of Korean language in computer chatting rooms, which are relevant to hypertext, as a part of preparing the contents and methods of hypertext expression education\n\n First, because of the 'anonymity' of computer chatting rooms, people can expression their feelings without concealing. On the other hand, anonymity causes flaming and undermines linguistic morality. Conversations in computer chatting rooms occur through direct feedback using 'interaction' Because many meetings and conversations in computer chatting rooms are improvised, they are more impersonal than sincere, and politeness and rules necessary for conversation are often ignored Language in chatting rooms is exchanged by texts, but they contain oral elements, which are expressed through the mouth, as well as non-oral elements That is, the language. is mixed with oral words and textual words.\n\n Telecommunication language is used based on strategies for preserving efficiency, those for preserving expressivity, and psychological factors. Although some view it negatively saying that it destroys linguistic norms, its positive side of creative utilization of texts is not negligible, The reason that such telecommunication language is perceived negatively is its influence on everyday language. Concerning the influence, further research is required.\n\n Research related to the education of expression in computer chatting rooms is mostly focused on telecommunication language. As the present researcher investigated, in addition, contents in Writing and Korean Language Life in the 7th Education Curricula also deal mainly with the destruction of linguistic norms by telecommunication language and problems in linguistic morality.\n\n Thus, the researcher proposes plans for the education of Korean Language related to expressions in computer chatting rooms as follows. ① It is necessary to understand telecommunication language positively, regarding it as a social dialect. ② In a sense, abbreviations, emoticons (emotion icons) and symbolic words are language creation, which are necessary for efficient conversations in communication. ③ Except the examples presented in ②, telecommunication language should not be transferred to everyday language, and for this students must learn grammar intensively and have an ability to distinguish virtual worlds from the real world ④ Students must be given not only theoretic education on the characteristics of oral and textual words rot also one related to conversational expressions in computer chatting rooms in 'Speech' and 'Narration' classes.
this paper I will discuss a framework for semantics which allows us to record truth-conditional and compositional analyses as dependency-style corpus annotations in a direct and fine-grained fashion. This method eliminates the need for a semantic representation formalism by decomposing semantic information into simple statements about (word or morpheme) tokens. A collection of such data would form a new kind of linguistic treebank. The main purpose of this article is to show that the present approach makes it possible to combine formal semantics and corpus-oriented study of language use in new and interesting ways. The methodology of this framework, which I call Token Dependency Semantics (TDS, Dahllf [4]), is in several respects different from the common one(s) in traditional formal semantics. TDS nevertheless delivers a fairly conventional (but ontologically restrained) analysis of truth-conditional meaning
The paper deals with a special kind of comparative without an overt secundum comparationis, as exemplified by, say, Pale su jače kiše 'stronger rains have fallen', which is freely used in Serbian; it is called, according to the grammatical tradition, absolute comparative. Attention has been drawn, while attempts were made at revealing the crucial features of the absolute comparative, to the fact that it is marked for its "amplified extension" on the scale of gradation; this shows that the comparative in the kind of use now under consideration has not lost its nature of an instrument of comparison. The grammatical structure in question has also been characterized as displaying sui generis semantic indefiniteness: the point is that the lack of the second object of comparison makes the scope of application of a given feature on the gradation scale rather fuzzy. The first part of the paper presents the distribution of the absolute comparative within the Slavonic linguistic area; the presentation is based on the existing grammars of particular Slavonic languages (however, Bulgarian and Macedonian have not been accounted for; the reason was that these languages have been "balcanized"). The evidence supplied by grammars allows us to distinguish, within the Slavonic area, two zones: the zone of marginal use of the form in question (Russian) or its limited use (Polish), and the zone of its active use, including Slovak, Czech, Sorbian, Slovenian and Serbian. The second part of the work describes the contrast between the situation in Serbian, on the one hand, and the situation in Polish, on the other: the focus is on the distinct divergence of the two languages in terms of textual distribution, frequency of occurrence and stylistic characteristics of the investigated structure. On the basis of the materials of bilateral translations of belletristic works, as well as those of the Serbian journalistic texts (as appearing in Internet), selected types of translational equivalences of the Serbian absolute comparative in Polish texts have been discussed; these are: the basic adjective in the positive, the negated antonym of the source adjective, and the construction "co + adjective in the comparative degree". In the last part of the article some selected differences concerning the use of the absolute comparative in Serbian and Polish journalistic texts have been pointed out. As shown in the course of the analysis, the Polish journalistic style tends to express sharp appraisals and distinct evaluations. This is particularly evident in isolated elements of press, such as titles, notices, advertising slogans. As a result, the absolute comparative, with its considerable degree of indefiniteness, appears to be less appropriate here. In contrary to this, the Serbian linguistic norm admits of a milder form of utterance, it admits of formulating less categorical judgments, even in journalistic style; this enhances the use of the absolute comparative which is well anchored both in the grammatical system and in linguistic awareness of the users of Serbian.
Abstract This paper describes a framework for building story traces (compact global views of a narrative) and story projections (selections of key elements of a narrative) and their applications in text understanding and classification. Word and sense properties are extracted from documents using the WordNet lexical database enhanced with Prolog inference rules and a number of lexical transformations. Inference rules are based on navigation in various WordNet relation chains (hypernyms, meronyms, entailment and causality links, etc.) and derived relations expressed as Boolean combinations of node and edge properties used to direct the navigation. The resulting abstract story traces provide a compact view of the underlying narrative's key content elements and a means for automated indexing and classification of text collections. Ontology driven projections act as a kind of “semantic lenses” and provide a means to select a subset of a narrative whose key sense elements are subsumed by a set of concepts, predicates and properties expressing the focus of interest of a user. Finally, we discuss applications of these techniques in text understanding, classification of text collections and answering questions about a text.
(2002) Br J Psychiatry 180, 523; Turkington D, Kingdon D, Turner T.. Effectiveness of a brief cognitive-behavioural therapy intervention in the treatment of schizophrenia..;.:. –7. [OpenUrl][1][Abstract/FREE Full Text][2] QUESTION: In patients with schizophrenia in secondary care settings, does cognitive behavioural therapy (CBT) delivered by community psychiatric nurses (CPNs) improve symptoms? Randomised {allocation concealed*}†, unblinded,* controlled trial with 2–3 months of follow up. 6 centres in the UK (Belfast, Glasgow, Hackney, Newcastle, Southampton, and Swansea). 422 patients who were 18–65 years of age (mean age 40 y, 77% men) and were receiving treatment from psychiatric secondary care services. Exclusion criteria were need for inpatient care or intensive home treatment, primary diagnosis of drug or alcohol dependence, organic brain disease, or learning disability that could affect rating. Follow up was 84%. Patients were allocated to CBT … [1]: {openurl}?query=rft.jtitle%253DThe%2BBritish%2BJournal%2Bof%2BPsychiatry%26rft.stitle%253DBr.%2BJ.%2BPsychiatry%26rft.issn%253D0007-1250%26rft.aulast%253DTURKINGTON%26rft.auinit1%253DD.%26rft.volume%253D180%26rft.issue%253D6%26rft.spage%253D523%26rft.epage%253D527%26rft.atitle%253DEffectiveness%2Bof%2Ba%2Bbrief%2Bcognitive--behavioural%2Btherapy%2Bintervention%2Bin%2Bthe%2Btreatment%2Bof%2Bschizophrenia%26rft_id%253Dinfo%253Adoi%252F10.1192%252Fbjp.180.6.523%26rft_id%253Dinfo%253Apmid%252F12042231%26rft.genre%253Darticle%26rft_val_fmt%253Dinfo%253Aofi%252Ffmt%253Akev%253Amtx%253Ajournal%26ctx_ver%253DZ39.88-2004%26url_ver%253DZ39.88-2004%26url_ctx_fmt%253Dinfo%253Aofi%252Ffmt%253Akev%253Amtx%253Actx [2]: /lookup/ijlink?linkType=ABST&journalCode=bjprcpsych&resid=180/6/523&atom=%2Febmed%2F8%2F1%2F23.atom
The purpose of present study was to investigate the effect of school and classroom images on adjustment to the school among children. In Study 1, two scales were constructed to assess school and classroom images in elementary school children. In the school image scale, factor analysis yielded 4 factors: positive exterior, negative exterior, dominance, and safeguard. In the classroom image scale, factor analysis yielded 4 factors: relief, crowdedness, dominance, and cheerfulness. These findings suggest that schools and classrooms allow children to project their thought, feeling, conflicts, and frames of mind. In Study 2, Multiple-regression analyses were performed on the variables of school and classroom image ratings, using school and classroom image ratings as independent variables and school moral test ratings as dependent variables. The results shows that feelings of being dominated have an effect on children adjustment to the school.
The thesis is titled Mario Monteleone Electronic Dictionaries and Lexicography. Uses language with lexical databases provided concerning the relationship (and potential) between traditional lexicography and computational linguistics. In particular, Mr. Monteleone is trying to establish whether some of the lexicography analytical limits can be exceeded through the application of research methods in computational linguistics. To this end, Mr. Monteleone is conducting a detailed analysis of the two disciplines, of their traditional instruments, ie dictionaries paper and electronic dictionaries, in terms of their construction methods and their application areas. The result is a handbook for traditional lexicographers and computer focused on the identification and classification of linguistic data for the accomplishments of paper and electronic dictionaries.
The aim of this article is to show that the role of legal terminology in juridical discourse can adequately be determined in a framework of a "juridical textwork" model only. This implies to analyse their role under pragmatic and intertextual aspects. lt will be contended that neither semantics nor the wording of the law can be employed as a starting point for an adequate interpretation of legal language in public discourse. Of crucial importance is rather the question of how the legal text and social reality can combine to constitute the legal norm (which is more than just the legal text). This is illustrated with the example of German court decisions on the issue of sit-ins that were organised by the peace movement in the l970s and l980s to block the access to American army bases. lt will be demonstrated that in the process of putting the coercion law in concrete normative terms through different courts, specific legal terms are semanticly modified and adjusted to specific language use, hence constituting "semantic battles" and linguistic norm conflicts in the juridical discourse.
This paper describes the use of clustering at three stages within a larger research effort to identify semantic frames used in English automatically. The first of two tasks within this effort has been the identification of sets of semantically related verb senses that invoke a common semantic frame. Within this task, clustering has been used both to build sets of verb senses with the potential of invoking a common semantic frame and then to merge sets with a high degree of overlap. The paper is organized as follows: Section 2 introduces frame semantics. Section 3 outlines the methodology used to identify sets of semantically related verb senses that invoke a common semantic frame, while section 4 presents the specific clustering algorithm used within that process. Section 5 discusses the use of this clustering algorithm for the identification of semantically related verbs in two machine-readable lexical resources: the machine-readable version of the Longman Dictionary of Contemporary English (LDOCE, 1978 edition) and WordNet, an online lexical database (http://www.cogsci.princeton.edu/-wn; version 1.7.1 has been used for the work reported here). Section 6 presents the use of clustering to merge overlapping sets of verb senses formed in previous steps. Section 7 discusses the results of these clusterings, paying particular attention to the effect of LDOCE's restricted defining vocabulary on the clustering process.
Choosing the statistical model is the key problem in statistical parsing. Statistical model lies in the core of NLP parsing. This paper investigates 4 primary statistical parsing models, namely PCFG, history-based model, cascading parsing model and head-driven parsing model, and compares their performances in a 10000 Chinese treebank. The analysis based on the experiment were shown in the paper. The comparative study of these models can be exploited to build the practical and effective Chinese parser.
The aim of this article is to show that the role of legal terminology in juridical discourse can adequately be determined in a framework of a juridical textwork model only. This implies to analyse their role under pragmatic and intertextual aspects. It will be contended that neither semantics nor the wording of the law can be employed as a starting point for an adequate interpretation of legal language in public discourse. Of crucial importance is rather the question or how the legal text and social reality can combine to constitute the legal norm (which is more than just the legal text). This is illustrated with the example of German court decisions on the issue of sit-ins that were organised by the peace movement in the 1970s and 1980s to block the access to American army bases. It will be demonstrated that in the process of purling the coercion law in concrete normative terms through different courts, specific legal terms are semanticly modified and adjusted to specific language use, hence constituting semantic battles and linguistic norm conflicts in the juridical discourse.
Computer games and the technologies marketed to support them provide unique resources for psychological research. In contrast to the sterility, simplicity, and artificiality that characterizes many cognitive tests, game-like tasks can be complex, ecologically valid, and even fun. In the present paper, the history of psychological research with video games is reviewed, and several thematic benefits of this paradigm are identified. These benefits, as well as the possible pitfalls of research with computer game technology and game-like tasks, are illustrated with data from comparative and cognitive investigations.
Psychology has to deal with many interacting variables. The analyses usually used to uncover such relationships have many constraints that limit their utility. We briefly discuss these and describe recent work that uses genetic programming to evolve equations to combine variables in nonlinear ways in a number of different domains. We focus on four studies of interactions from lexical access experiments and psychometric problems. In all cases, genetic programming described nonlinear combinations of items in a manner that was subsequently independently verified. We discuss the general implications of genetic programming and related computational methods for multivariate problems in psychology.
Acronyms are a very dynamic area of the lexicon of many languages. A hybrid, modular methodology for the acquisition of acronyms is presented, which uses an existing acronym-expansion matching component, and machine learning in two separate phases for the identification of long-distance acronym definition patterns.The resulting system, using Support Vector Machines (SVM) is trained on 600 news stories from the Wall Street Journal component of the Penn Treebank corpus using a number of lexical, syntactic, and acronym-expansion matching features. Statistical cooccurrence information for acronym-expansion pairs is extracted from search engine hit counts.The system achieves Fβ=1=92.38% on 400 news stories from the same source and has good asymptotic efficiency, making it adequate for the automatic extraction of acronyms even from noisy sources, such as newspaper text.
Among the most consistent findings in the warnings literature is the so-called 'familiarity effect.' Research has shown that the more familiar an individual is with a product or situation the less likely he or she is to notice, read, recall, or comply with hazard communications. The effect has been found across numerous product types and situations using various operational definitions of familiarity and measures of warning effectiveness. However, research has also shown that subjective familiarity ratings are not highly correlated with actual product experience. Thus, individuals must be capable of developing a false or exaggerated sense of familiarity. One possible source of this exaggerated familiarity is exposure to product advertising.\n\nThree experiments were conducted to investigate whether the familiarity effect can be produced from exposure to product advertising. The relationships between advertising exposure and perceived familiarity and between perceived familiarity, perceived safety and warning effectiveness were examined. Experiment 1 explored participants' attitudes and beliefs about well-known and obscure brands of household, consumer products and sought to determine how past, direct product experience influences those attitudes and beliefs. Experiments 2 and 3 examined how the number of advertising exposures and the safety-related content of advertisements influence attitudes and beliefs about the advertised products and the effectiveness of on product warnings.\n\nResults of Experiment 1 revealed that past experience can not fully explain consumers' attitudes and beliefs about household, consumer products. Experiments 2 and 3 showed that advertising influences perceived product familiarity and knowledge. While there was a trend of greater perceived safety with increased ad exposures, the effect was not significant. No effects of advertising on warning recall were found. Implications for the design of product advertisements and product packaging as well as directions for future research are discussed.
Recovering the semantics, or the meaning, expressed using natural languages is one of the central goals of the field of natural language processing. The task is challenging due to the large number of ambiguities present in natural languages, such as part-of-speech assignments, structural dependencies, and word sense ambiguity. Reliably resolving these ambiguities would allow computers to gain access to the knowledge represented using natural languages, the format most commonly used in human communications. In this thesis we approach this complex problem by first decomposing it into four smaller problems, or subtasks: part of speech (POS) tagging, word sense disambiguation (WSD), chunking, and parsing. With this decomposition each subtask captures a subset of the ambiguities, thus simplifying the problem and facilitating the application of exact algorithms to maximize accuracy. We then integrate the decisions from the subtasks to form the most plausible interpretation, in contrast to top-down modeling that is prone to error propagation. We first apply machine learning algorithms to automatically train probabilistic models based on annotated training corpus. We select a powerful probability model, maximum entropy, to incorporate diverse contexts systematically. We then capture the dependencies between the words within sentences using Bayesian networks, with which we compute disambiguation decisions using an exact inferencing algorithm. To share information between subtasks, we introduce a new structural representation, called Cores-and-Modifiers, to succinctly describe structural features in improving both POS tagging and WSD. We also identify semantic contexts based on the WordNet lexical database to improve both chunking and parsing accuracy. To form the overall interpretation across these subtasks, we introduce an integrative process, instead of a top-down model. Because the entire model is probabilistic, it enables the systematic re-integration of the diverse decisions from the subtasks based on their probabilities. To further improve this integration, we present the concept that the most uncertain decisions are the most error-prone, and thus their alternatives should also be examined. This process, named Most-probable Hypotheses Evaluation, selectively examines a small set of alternate hypotheses to better determine the most plausible interpretation across subtasks, instead of as disparate decisions. The resulting model, named Integrative, Probabilistic Natural-language Parser and Interpreter (IPNPI), integrates the four subtasks to form the most plausible interpretations. The IPNPI model is evaluated by its accuracy on each of the four subtasks using standardized procedures, and we show that it improves the state-of-the-art accuracy in POS tagging, word sense disambiguation, chunking, and parsing. The synthesis of our separate-but-integrative approach is that the IPNPI model is able to accurately and efficiently resolve natural language ambiguities, by producing the most plausible interpretations across part of speech assignment, word sense distinction, phrasal identification, and structural dependencies.
Crosslinguistically vocatives are an underexplored linguistic phenomenon and in different languages they can be highly idiosyncratic and complex (Levinson, 1987, p.71). Therefore, the problem, which is discussed in this paper, is not a language-specific one, in spite of the fact that most of the languages have their own repositories for marking the role of the addressee in the communicative utterances. In our opinion this linguistic phenomenon needs its adequate treatment in HPSG because of three main reasons: 
 
 The vocative is supposed to be present on two levels: syntax and pragmatics. Therefore it needs more elaborate interpretation on the interface side, which, in HPSG, is more developed for morphology/syntax and syntax/semantics than syntax/pragmatics. Note that a challenge for the theory is the semantic weight of the vocatives with respect to the head sentence. 
 It will be useful for HPSG-oriented implementations, especially treebanks and dialogue systems. 
 On prosodic grounds the vocatives are often viewed as being 'side or extended parts' of the sentence and therefore - very close to the parenthetical constructions. From our point of view, both phenomena are pragmatic and hence, the treatment of vocative, presented here, could be generalized to cover other phenomena of pragmatic nature. 
 
 In our work the vocatives are viewed through the possibility of the integration/separation of their pragmatic, syntactic and semantic properties.
In den letzten Jahren ist die Zahl der verfgbaren linguistisch annotierten Korpora stndig gewachsen. Zu den bekanntesten gehren das Brown-Korpus, das Susanne-Korpus, die Penn-Treebank, das Negra-Korpus, das Tiger-Korpus und die im Zusam-