Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
Language and language processing technologies are an important basis for research in the field of information science. The development of linguistic databases, such as electronic dictionaries (machine-readable/tractable dictionaries), contributes to the development of stable language processing technologies. The information research potential of the electronic dictionary greatly broadens and expands the field of information science. The fields of knowledge processing and information science greatly overlap. The building of very large knowledge bases, which will fully stabilize knowledge processing technologies, can be most effectively approached by considering/utilizing research in linguistic knowledge in combination with research in the technology of language processing. This paper will discuss the EDR Electronic Dictionary in relation to the development of very large knowledge bases.
The Robert Electronique is the CD-ROM version of the nine volume Grand Robert, roughly the French equivalent of the OED. This article outlines a project to produce some learning materials using this lexical database. It describes various types of exercises ranging from the semantic to the stylistic. Most of the exercises can be completed on screen within a wordprocessing application and can be done by a student working independently. The activities exploit as far as possible features specific to the CD-ROM version of the dictionary.
One experiment compared the effect of elaboration on enacted and non-enacted events. The commands were either presented in a basic form (e.g., "wave your hands") or in an enriched form. The commands were enriched by adding statements to the commands of how to perform the actions (e.g., "wave your hands as a conductor"). Free- and cued-recall data showed elaboration to have a dissociative effect on enacted and non-enacted events. Memory for the non-enacted events benefited from enrichment, whereas simple enacted events were remembered to a higher extent than complex enacted events. Lack of benefit from elaboration on memory of enacted events is suggested to be due to enactment leading to a sufficient degree of item-specific processing, and a negative effect of elaboration is suggested to occur when the way of manipulating item complexity decreases the familiarity of the actions. Familiarity ratings of the items by two independent groups of subjects supported this interpretation.
The field of natural language processing (NLP) has seen a dramatic shift in both research direction and methodology in the past several years. In the past, most work in computational linguistics tended to focus on purely symbolic methods. Recently, more and more work is shifting toward hybrid methods that combine new empirical corpus-based methods, including the use of probabilistic and information-theoretic techniques, with traditional symbolic methods. This work is made possible by the recent availability of linguistic databases that add rich linguistic annotation to corpora of natural language text. Already, these methods have led to a dramatic improvement in the performance of a variety of NLP systems with similar improvement likely in the coming years. This paper focuses on these trends, surveying in particular three areas of recent progress: part-of-speech tagging, stochastic parsing, and lexical semantics.
Abstract By comparing translations and retranslations of several children's books into Hebrew done over a span of 70 years we try to find out what linguistic and translational norms prevailed at different periods and what changes occurred in these norms, in a framework of the changing historical, cultural and linguistic situation. In recent years there has been a growing tension between the acceptability of the "commissioners" and that of the "customer"—the child. This study found that recent retranslations tend to lower the high literary style customary in previous translations and comply with up-to-date linguistic norms. This concurs with a tendency to put "readability" as a central issue.
The prevalence of the use of teams in a variety of occupations and environments has increased the importance of investigating the processes involved in their performance. However, in the past, there have been few methodologies available for the investigation of team performance. The present manuscript attempts to contribute to this area of research by describing the rationale underlying the use of computer-based simulations in research on team performance. This is followed by a review of the networked simulations that are currently being used in team-performance research. This review emphasizes the capabilities provided by the networks and the types of research concerns for which they are effective. Finally, the application of this technology to the broader study of group performance is discussed.
We set out to develop a computer-assisted finger-tapping task (the T3) that would measure motor speed much like the Reitan test, but that would also measure endurance. Data were collected for a convenience sample on both the T3 and the Reitan finger-tapping test. Moderate and significant correlations were obtained between the T3 and the Reitan test for both hands. Mean scores for the first 50 sec of the T3 were approximately 0.15 taps greater than the mean Reitan score for both the preferred and the nonpreferred hands, while the mean scores for the full 2 min of the T3 were 1.52 taps less than those of the Reitan test for the preferred hand, and 1.32 taps less for the nonpreferred hand. The mean for the last 40 sec with the preferred hand averaged 3.93 taps (7.62%) slower than for the first 40 sec, whereas for the nonpreferred hand, the difference was 5.12 taps (11.15%). These results are consistent with our intent to develop measures of (1) relatively pure motor speed (the first 50 sec of the T3); (2) motor speed combined with endurance (the full 2 min of the T3); and (3) finger endurance (the first 40 sec compared with the last 40 sec of the T3).
Many languages make use of word-formation devices to allow speakers or writers to create new words when the existing vocabulary proves inadequate. In this paper we consider how these devices can be expressed formally, allowing them to be used in word- and sentence-generation, for dictionary expansion, and the like. The paper begins with some typical word-formation rules drawn mostly from French. Attention is drawn to some features of these rules which must be captured in any formal representation. The formal representation of a basic lexical transformation is presented in some detail, along with a number of examples. A computer implementation of the transformation system is described, together with a range of applications. A discussion of static and dynamic generation leads to the concept of an inverted transformation.
The article gives a general introduction to the form and function of the TEI header, points out some of the reasoning of the Text Documentation Committee that went into its design, and discusses some of its limitations. The TEI header's major strength is that it gives encoders the ability to document the electronic text itself, its source, its encoding principles, revisions, and characteristics of the text in an interchange format. Its bibliographical descriptions can be loaded into standard remote bibliographic databases, which should make electronic texts as easy to find for researchers as texts in other media, including print. Its major weakness is that it does not yet provide the ability for retrieval across texts in a networked environment, which users may want now or in the future.
In this paper, a method for indexing cross-language databases for conceptual query matching is presented. Two languages (Greek and English) are combined by appending a small portion of documents from one language to the identical documents in the other language. The proposed merging strategy duplicates less than 7% of the entire database (made up of different translations of the Gospels). Previous strategies duplicated up to 34% of the initial database in order to perform the merger. The proposed method retrieves a larger number of relevant documents for both languages with higher cosine rankings when Latent Semantic Indexing (LSI) is employed. Using the proposed merge strategies, LSI is shown to be effective in retrieving documents from either language (Greek or English) without requiring any translation of a user's query. An effective Bible search product needs to allow the use of natural language for searching (queries). LSI enables the user to form queries with using natural expressions in the user's own native language. The merging strategy proposed in this study enables LSI to retrieve relevant documents effectively using a minimum of the database in a foreign language.
Small groups are called upon to make important policy decisions under a wide variety of procedural constraints. ACPE is a flexible, computerized system for conducting small-group voting experiments. It permits researchers to examine the impact of electronic communication on group deliberation and choice. The system runs under a variety of different personal computer networks and is designed to permit the specification of voting rules, communication, and group sizes. The system also facilitates the study of group process by tracking all messages sent and votes taken. An experiment in which the system was used is briefly described.
This paper focusses on the types of questions that are raised in the encoding of historical documents. Using the example of a 17th century Scottish Sasine, the authors show how TEI-based encoding can produce a text which will be of major value to a variety of future historical researchers. Firstly, they show how to produce a machine-readable transcription which would be comprehensible to a word-processor as a text stream filled with print and formatting instructions; to a text analysis package as compilation of named text segments of some known structure; and to a statistical package as a set of observations each of which comprises a number of defined and named variables. Secondly, they make provision for a machine-readable transcription where the encoder's research agenda and assumptions are reversible or alterable by secondary analysts who will have access to a maximum amount of information contained in the original source.
Many aspects of the guidelines of the Text Encoding Initiative (TEI) are applicable to corpora and text collections, and to the texts that these contain. As the first large corpus developed using mark-up conforming to the guidelines, the British National Corpus (BNC) is a test-bed for many TEI-developed mechanisms. This is particularly true in the case of the TEI header, which has three intended applications — to describe a corpus, to describe an individual text, and as a free-standing bibliographic record — all of them used by the BNC. This paper describes the application of the TEI header to the BNC. It is intended that this information should, through a description of experience on a practical project, serve as a guide for those wishing to use TEI headers in the documentation and management of other corpora and collections of texts.
In this paper we show, for the first time, how Radial Basis Function (RBF) network techniques can be used to explore questions surrounding authorship of historic documents. The paper illustrates the technical and practical aspects of RBF's, using data extracted from works written in the early 17th century by William Shakespeare and his contemporary John Fletcher. We also present benchmark comparisons with other standard techniques for contrast and comparison.
There are many ways in which to estimate thresholds from psychometric functions. However, almost nothing is known about the relationships between these estimates. In the present experiment, Monte Carlo techniques were used to compare psychometric thresholds obtained using six methods. Three psychometric functions were simulated using Naka-Rushton and Weibull functions and a probit/logit function combination. Thresholds were estimated using probit, logit, and normit analyses and least-squares regressions of untransformed orz-score and logit-transformed probabilities versus stimulus strength. Histograms were derived from 100 thresholds using each of the six methods for various sampling strategies of each psychometric function. Thresholds from probit, logit, and normit analyses were remarkably similar. Thresholds fromz-score- and logit-transformed regressions were more variable, and linear regression produced biased threshold estimates under some circumstances. Considering the similarity of thresholds, the speed of computation, and the ease of implementation, logit and normit analyses provide effective alternatives to the current “gold standard”—probit analysis—for the estimation of psychometric thresholds.
In this paper, we concentrate on justifying the decisions we made in developing the TEI recommendations for feature structure markup. The first four sections of this paper present the justification for the recommended treatment of feature structures, of features and their values, and of combinations of features or values and of alternations and negations of features and their values. Section 5 departs briefly from the linguistic focus to argue that the markup scheme developed for feature structures is in fact a general-purpose mechanism that can be used for a wide range of applications. Section 6 describes an auxiliary document called a “feature system declaration” that is used to document and validate a system of feature-structure markup. The seventh and final section illustrates the use of the recommended markup scheme with two examples, lexical tagging and interlinear text analysis.
There is a great deal of variation in the encoding of spoken texts in electronic form, both with respect to the types of features represented and the way particular features are rendered. This paper surveys problems in the electronic representation of speech and presents the solutions proposed by the Text Encoding Initiative. The special tags needed for the encoding of spoken texts are discussed, including a mechanism for temporal alignment. Further work is needed on phonological aspects, parallel representation, and on the development of software which connects the systematic underlying representation with a workable format for input and display.
focuses on certain differences between Standard English (SE) and the African-American Vernacular English (AAVE) spoken in many inner-city and rural American communities / presents 2 points that are crucial for the treatment of languages that differ much more widely / 1st, vernacular languages—the ways that ordinary people ordinarily talk—are not 'ungrammatical' or otherwise imperfect approximations to standard or literary linguistic norms / 2nd, small changes in an abstract grammatical system may produce complex patterns of change on the surface of the language, magnifying the apparent differences (PsycINFO Database Record (c) 2016 APA, all rights reserved), no part of cognitive science illustrates the problem of abstract inference better than the interpretation of linguistic zeroes: the absence of the very behavior that we have come to observe / [discuss] the interpretation of such linguistic zeroes / engage a particular problem that has been the center of much linguistic research: t)
Natural language processing will grow into a vital industrial technology in the next five to 10 years. But this growth depends on the development of large linguistic databases that capture natural language phenomena [1, 2]. Another important theme for future work is development of large knowledge bases that are shared widely by different groups. One promising approach to such knowledge bases draws on natural language processing and linguistic knowledge. This article describes the EDR Electronic Dictionary [3], which seeks to provide a foundation for linguistic databases, and explains the relation of electronic dictionaries to very large knowledge bases.
Another current major issue in lexical semantics is the definition and the construction of real-size lexical databases that will be used by parsers and generators in conjunction with a grammatical system. Word meaning, terminological knowledge representation and extraction of knowledge in machine readable dictionaries are the main topics addressed. They really represent the backbone of a lexical semantics knowledge base construction.
Parsing is often seen as a combinatorial problem. It is not due to the properties of the natural languages, but due to the parsing strategies. This paper investigates a Constrained Grammar extracted from a Treebank and applies it in a non-combinatorial partial parser. This parser is a simpler version of a chunking-and-raising parser. The chunking and raising actions can be done in linear time. The short-term goal of this research is to help the development of a partially bracketed corpus, i.e., a simpler version of a treebank. The long-term goal is to provide high level linguistic constraints for many natural language applications. 1
Lexical Collocations are frequently occurring word pairs in natural language whose presence are not always predictable by their usage. These collocations are used by native speakers of a language almost without thought; yet they must be learned by non-native speakers of that language. A native speaker of English may drink strong coffee while a non-native speaker may say either $\sp{*}$powerful coffee or $\sp{*}$sturdy coffee. Collocations tend to vary among languages and topic domains. Unfortunately, the task of correctly identifying lexical collocations, even by native speakers of that language, has been shown to be very difficult. Computer systems that translate natural languages, or Machine Translation (MT) systems, need to know about lexical collocation information in order to produce natural sounding or colloquially proper text. Natural Language Generation (NLG) is a component of an MT system which automatically produces natural sounding text in a particular target language given a language-independent meaning as input. This dissertation will demonstrate how to automatically locate and extract lexical collocations from machine-readable text for use within an MT system's NLG component. A lexical-semantic and statistical approach is adopted for the location and extraction of lexical collocations. For this approach, a computational definition is provided for lexical collocations which demonstrates that: (1) they occur as adjacent word pairs; (2) they occur more often than would be expected by chance; and (3) they comprise words for which neither word may be substituted by a synonym or hyponym. Potential collocations comprising certain adjacent part-of-speech tags are extracted from text. An on-line thesaurus and lexical database of word classes are queried for synonyms and hyponyms, respectively, for each potential collocation. These queries create potential challenger pairs, such as strong java and powerful coffee. A substitution procedure is then applied to determine if any of these challenging word pairs occur more frequently than the potential collocation. The VERIFY lexical collocation extraction system has been implemented incorporating these ideas. Results to date have been positive: using lexical-semantic knowledge, i.e., synonymy and hyponymy, within a lexical collocation extraction system outperforms a system using purely statistical knowledge. In order to compare system output to human judgments of training data, a training component was also incorporated into VERIFY. This component is able to adapt to new data. Overall system performance, measured by Recall and Precision scores, was shown to improve using this component. In order to provide a more flexible system given a user's application, a weighting mechanism was used to produce a range of Recall and Precision scores. These weights can be 'adjusted' to optimize system performance. The use of lexical-semantic knowledge has advanced the state of the art for lexical collocation extraction beyond traditional statistical approaches. Incorporation of a training component within an extraction system provides the capability of adapting to any changes within the data. Controlling overall system performance through the use of a weighting mechanism provides flexibility to the user of an extraction system. And, in an experiment to compare VERIFY'S performance to that of human performance on a particular set of data, it was shown that VERIFY outperforms humans in both Recall and Precision.
In the 1980s the dominant framework of MT was essentially ‘rule‐based’, e.g. the linguistics‐based approaches of Ariane, METAL, Eurotra, etc.; or the knowledge‐based approaches at Carnegie Mellon University and elsewhere. New approaches of the 1990s are based on large text corpora, the alignment of bilingual texts, the use of statistical methods and the use of parallel corpora for ‘example‐based’ translation. The problems of building large monolingual and bilingual lexical databases and of generating good quality output have come to the fore. In the past most systems were intended to be general‐purpose; now most are designed for specialized applications, e.g. restricted to controlled languages, to a sublanguage or to a specific domain, to a particular organization or to a particular user‐type. In addition, the field is widening with research under way on speech translation, on systems for monolingual users not knowing target languages, on systems for multilingual generation directly from structured databases, and in general for uses other than those traditionally associated with translation services.
An extragrammatical sentence is what a normal parser fails to analyze. It is important to recover it using only syntactic information although results of recovery are better if semantic factors are considered. A general algorithm for least-errors recognition, which is based only on syntactic information, was proposed by G. Lyon to deal with the extragrammaticality. We extended this algorithm to recover extragrammatical sentence into grammatical one in running text. Our robust parser with recovery mechanism -- extended general algorithm for least-errors recognition -- can be easily scaled up and modified because it utilize only syntactic information. To upgrade this robust parser we proposed heuristics through the analysis on the Penn treebank corpus. The experimental result shows 68% ¸ 77% accuracy in error recovery. 1 Introduction Extragrammatical sentences include patently ungrammatical constructions as well as utterances that may be grammatically acceptable but are beyond the synta...
Syntactic natural language parsers have shown themselves to be inadequate for processing highly-ambiguous large-vocabulary text, as is evidenced by their poor performance on domains like the Wall Street Journal, and by the movement away from parsing-based approaches to text-processing in general. In this paper, I describe SPATTER, a statistical parser based on decision-tree learning techniques which constructs a complete parse for every sentence and achieves accuracy rates far better than any published result. This work is based on the following premises: (1) grammars are too complex and detailed to develop manually for most interesting domains; (2) parsing models must rely heavily on lexical and contextual information to analyze sentences accurately; and (3) existing n-gram modeling techniques are inadequate for parsing models. In experiments comparing SPATTER with IBM's computer manuals parser, SPATTER significantly outperforms the grammar-based parser. Evaluating SPATTER against the Penn Treebank Wall Street Journal corpus using the PARSEVAL measures, SPATTER achieves 86% precision, 86% recall, and 1.3 crossing brackets per sentence for sentences of 40 words or less, and 91% precision, 90% recall, and 0.5 crossing brackets for sentences between 10 and 20 words in length.
espanolEn castellano, la silaba parece actuar como un elemento prelexico-fonologico de relacion con el nivel lexico. Su mayor o menor frecuencia determina la cantidad de palabras que se activaran en el nivel lexico. Esta cualidad de restriccion lexica nos ha sugerido la necesidad de elaborar un estudio normativo en el cual los sujetos evocaban palabras de 2 y 3 silabas a partir de una inicial dada que despues pueden ser utilizados como base para diversos estudios experimentales. Se produjeron asi 130 conjuntos de candidatos competidores lexicos (ccl), que proporcionan informacion sobre dos aspectos fundamentales: su tamano (numero de candidatos lexicos) y la accesibilidad relativa de cada una de las salidas lexicas que las componen. EnglishThe syllable in Spanish could operate as a phonological and prelexical unit with relation to lexical level. The syllable frequency determines the number of words that will be activated at lexical level. This quality of accessibility constriction suggests the need to produce some candidate set norms. In this normative study, the subjects recover two and three syllable words from one initial syllable given. In this way, we obtained 130 sets of lexical competitor candidates (CCL), which provide information about two basic aspects: the set size (i.e. number of lexical candidates), and relative accessibility of each lexical output composing the set.
Abstract OUR objective in this book is to trace and illustrate the main changes that have taken place in the Russian language since the beginning of the twentieth century, and particularly those between 1917 and the late 1980s, a time which can be identified as the Soviet period, by now a reasonably complete stage in the history of Russia and the former Soviet Union. The period addressed in this book witnessed the unprecedented expansion of the standard Russian language, both stylistically and geographically. Around 1900, the range of functions of standard Russian increased, in some instances replacing Church Slavonic, in some instances with the emergence of new phenomena. Russian replaced Church Slavonic as the language of sermons (parallel to the maintenance of Church Slavonic as the language of liturgy); standard Russian emerged as the language of the courts (especially following the judicial reform of the 1860s), of political debate in the Russian Parliament (Duma), of growing business and trade, and as the language of poetry and prose read to large audiences from the stage of concert-halls and literary cabarets. Theatre, the most popular type of entertainment in Russia of the second half of the nineteenth century and at the beginning of the twentieth century, played a very important role in promoting a uniform language norm, especially in pronunciation; the role of theatre as the bearer of such a norm became clear by the mid 1Soos (Panov 1990: 94). In the twentieth century, radio, cinema, and television, gradually taking over from the theatre, gave the standard language the ability to travel, unknown in earlier periods; mass media, together with compulsory mass education, became much more effective than writing alone in acquiring new speakers for the emerging standard. These new speakers of the standard, in tum, had a profound effect on the language itself: never before had it borne the fingerprints of so many diverse speakers, many of whom were, following M. V. Panov’s characteriza tion, ‘new recruits to culture’ (Panov 1990: 18, 22).
EDBL est un base de donnees lexicale pour la langue basque. Cet article presente la conception et les caracteristiques principales de cette base de donnees, imaginee comme une base lexicographique generale pour le traitement automatique du basque. Le schema conceptuel de EDBL est explique au moyen du diagramme E/A etendu et de structures de traits. L'implementation de la base en tant que SGBD relationnel commercialisable et les problemes rencontres lors de cette implementation sont discutes
The relationship between positive and negative events and emotional well-being for depressed and nondepressed residents of a nursing home and congregate housing care facility was examined. For 30 consecutive working days, each of 79 participants was presented with the Philadelphia Geriatric Center Positive and Negative Affect rating scales. Events during the previous 24 hr were elicited by an open-ended format. Results indicated that variations in daily events (e.g., health, family, self-initiated, and social events) were related to residents' affect, and there was congruence between mood and event valence when the effects of psychopathology and residence were removed. Thus, regardless of diagnosis or residential setting, people's moods showed a relationship to the quality of daily events. Findings also indicated that ratings of residents' affect could be translated into audits for institutional quality.
Speech perceived on the basis of viewing a talker’s face affords less phonetic distinctiveness than acoustic speech. Effects of this reduced distinctiveness can be estimated in relation to the structure of the mental lexicon. Based on empirical measures of phonetic confusability, recoding rules can be defined for mapping fully specified phonological forms into lexical equivalence classes. For example, under the recoding rule that /b/ and /p/ are in the same phonemic equivalence class the words ‘‘bat’’ and ‘‘pat’’ map into the same lexical equivalence class. After applying a set of recoding rules to a large online lexical database, the resulting structure of the lexicon can then be studied quantitatively. One such measure of the recoding effects on the lexicon is percent information extracted (PIE) [D. M. Carter, Comput. Speech Lang. 2, 1–11 (1987)]. Lexical statistics describing the results of applying sets of recoding rules derived from analyses of visual-phonetic confusability to a 30 000-entry lexicon will be presented. Implications for the use of top-down lexical constraints in resolving bottom-up visual-phonetic ambiguity during lipreading will be discussed. [Work supported by NIH.]
The concept of class is studied by social psychologists as an important determinant in the development and expression of attitudes and behaviors (Schaefer, 1986). This dissertations explores existing perspectives of class including the theoretical treatments of experts and non-experts. The empirical use of the class concept is considered and past use is described as lacking some empirical basis. Study One serves to define the dimensions of the class construct as seen by non-experts. Thirty-two introductory psychology students rated eight class labels on fifty bi-polar adjectives. Multi-dimensional scaling revealed a two dimensional structure for class with ratings based on both a stereotypical hierarchical social class dimension and on a weaken ingroup bias evaluative dimension. Study Two serves to identify distinguishable class labels as well as differentiating characteristics of those labels. This allows for designation of characteristics or descriptions which are useful for social psychological research and have an empirical base. One hundred eighteen Introductory psychology students rated one class label on 52 bi-polar adjectives. Results reveal three distinct groupings, roughly corresponding to traditional upper, middle, and lower class delineation. However, working class and middle class are rated equivalently rather than as a dichotomy. The rating pattern replicates Study One by following either a traditional stereotypical delineation or self-identified class labels as more positive. In Study Three, class bias is explored using empirically designated descriptors by acquiring ratings for comparison between the two extreme class categories; upper and lower class. Other variables of interest are the valence of the description and presence or not of the class label. The question of class bias as a process distinct from race bias is explored by use of rating comparisons between class and race label-only stimuli. Responses from 246 introductory psychology student volunteers reveal that valence of the stimulus is a primary elicitor of reaction. They also indicate that class and an overt display of class membership via presence of a label affect ratings. Class bias is expressed overtly while racial bias is expressed subtlely.
Previous research (Pisoni and Garber, 1990; Garber and Pisoni, 1991) has demonstrated that subjective familiarity judgments for words are not differentially affected by the modality (visual or auditory) in which the words are presented, suggesting that subjects base their judgments on fairly abstract, modality-independent representations in memory. However, in a recent large scale study in Japanese (Amano etal., in press), markedly modality effects on familiarity ratings were observed. The current research further examines possible modality differences in subjective ratings and their implications for word recognition. Specially selected words were presented to subjects for frequency and recency judgments. In particular, subjects were asked how frequently (or recently) they READ, WROTE, HEARD, or SAID a given spoken or printed word. These ratings were then used to predict accuracy and processing times in auditory and visual lexical decision and naming tasks. Our results suggest modality dependence for some lexical representations, primarily for words that occur fairly rarely in the language. [Work supported by NIDCD.]
Adult children of alcoholics' (n = 68) perceptions of their relationships with parents were compared with those of a control sample (n = 37) to examine independent and joint influences of interpersonal status and affect on family dynamics. Visual metaphors for relationships using circle drawings and a status-affect rating scale from the Grasha-Ichiyama Psychological Size and Distance Scale were employed. Compared with the control group, adult children of alcoholics drew smaller circles to represent themselves, i.e., indicating less interpersonal status, only when assessing their relationships with their fathers. Analyses of status-affect ratings showed that the drawings of smaller circles reflected feeling less competent, i.e., having less personal knowledge and expertise, rather than perceptions of being submissive in the relationship. The distance drawn between the circles of adult children of alcoholics and their parents, i.e., psychological distance, was much larger than that of the control group. Ratings showed that perceptions of a negative emotional climate and submissiveness together accounted for 25% of the unique variance in predicting psychological distance. Perceptions of being submissive, however, were not associated with perceptions of psychological distance among adult children of nonalcoholic parents.
The present study examined factors hypothesized to influence mental health professionals' perceptions of dangerousness, predictions of violence, and decisions on patients' release. 120 mental health professionals employed in state mental hospitals were each given one of 12 patient profiles. The independent variables, manipulated within vignettes, were (a) violence history, (b) paranoid schizophrenia versus nonparanoid schizophrenia, and (c) perceived consequences in terms of liability and publicity. Type of schizophrenia did not affect ratings, but violence history of the predictee and perceived consequences to the predictor did significantly influence the ratings. Patients with actual violence histories were viewed by the subjects as having more potential for future violence, as being more globally dangerous, and as requiring a more secure placement than those with histories of threats of violence or no violence. Possible litigation following release led to a recommendation for more secure placement than did minimal legal consequences. Predictions of violence and decisions on hospital release were interpreted as dependent on both predictor and patient-related variables.
There are currently two philosophies for building grammars and parsers – Statistically induced grammars and Wide-coverage grammars. One way to combine the strengths of both approaches is to have a wide-coverage grammar with a heuristic component which is domain independent but whose contribution is tuned to particular domains. In this paper, we discuss a three-stage approach to disambiguation in the context of a lexicalized grammar, using a variety of domain independent heuristic techniques. We present a training algorithm which uses hand-bracketed treebank parses to set the weights of these heuristics. We compare the performance of our grammar against the performance of the IBM statistical grammar, using both untrained and trained weights for the heuristics. 1
This exploratory field study evaluated a bilingual computerized speech-recognition cellular telephone prototype of the Center for Epidemiological Studies—Depression scale (CES-D). Thirty Spanish and 22 English speakers completed both computer-telephone and face-to-face CES-D methods and an oral depression checklist in counterbalanced order. Both language groups reported high positive ratings for the computer-telephone method, with the English sample preferring the computer-telephone over the face-to-face method. In both samples, the computer-telephone method yielded high internal consistency estimates, strong alternate form reliabilities, and similar high correlations to the depression checklist. Both groups reported significantly elevated scores with the computer-telephone method, but total score variances for both methods did not differ. Computer-telephone limitations included occasional misrecognitions and template training constraints.
As a result of an analysis of the language of Jerzy Żuławski’s poetry, some devices have been selected which go beyond the linguistic norms common and universally accepted on the turn of the XIX and XX centuries. The selected linguistic devices can be divided into traditional, known in literature, and idiosyncratic, which make the language of the poems somewhat original. All these devices have been described and evaluated in chapters on phonetics, inflexion, syntax and lexis. Phonetic idiosyncracies are scarce and not many of them are typical of artistic language. Inflexion reflects the commonly accepted norms and patterns of style. Trying to make the poems original and their language individual, the poet used syntactic means more frequently than lexical ones, which is not typical of methods of stylization. The language of Żuławski’s poetry is „classical"; moderately permeated with traditional stylistic devices; slightly „poeticized”, archaized and individual. It lacks dialectal phrases and expressive devices characteristic of Modernist poets.
This paper discusses the basic design of the encoding scheme described by the Text Encoding Initiative'sGuidelines for Electronic Text Encoding and Interchange (TEI document number TEI P3, hereafter simplyP3 orthe Guidelines). It first reviews the basic design goals of the TEI project and their development during the course of the project. Next, it outlines some basic notions relevant for the design of any markup language and uses those notions to describe the basic structure of the TEI encoding scheme. It also describes briefly the “core” tag set defined in chapter 6 of P3, and the “default text structure” defined in chapter 7 of that work. The final section of the paper attempts an evaluation of P3 in the light of its original design goals, and outlines areas in which further work is still needed.
For a decade or so, Liszt thrilled and astounded audiences at a time when virtuosity (often as an end in itself) was the norm and the piano had rapidly evolved into a form recognisable as a close relative of the instrument we know today. During this period Liszt frequently performed hisGrandes Etudes (1838), which he had developed from his boyhoodEtude en 12 exercices (1826) and which he later revised and technically simplified asEtudes d'Exécution transcendante (1851). Although Liszt's own performances cannot be recreated, procedures for generating electronic realizations, which contain nuances of balance and tempo, are described. All three versions of the eighth of Liszt's set of 12 studies are used for illustration. Contrary to received opinion, it is argued that the 1838 version is more satisfying than the 1851 revision and that this is due to its formal structure.
This paper traces a progression of four computer-based methods for studying and fostering both the structure and the on-line development of knowledge. Each empirical technique employs ECHO, a connectionist model that instantiates the theory of explanatory coherence (TEC). First, verbal protocols of subjects’ reasonings were modeled post hoc. Next, ECHO predicted, a priori, subjects’ text-based believability ratings. Later, the bifurcation/bootstrapping method was developed to elicit and account for individuals’ background knowledge, while assessing intercoder reliability regarding ECHO simulations. Finally,Convince Me, our “reasoner’s workbench,” automated the explication both of subjects’ knowledge bases and of their belief assessments; theConvince Me software permits contrasts between the model’s predictions and subjects’ proposition-wise evaluations. These experimental systems enhance our understanding of the relationships among—and determinant features regarding—hypotheses, evidence, and the arguments that incorporate them.
The article contains a stylistic and linguistic analysis of Kasprowicz’s poems published in the period of Modernism. The following elements were selected from the texts of the poems: 1. deviations from the linguistic norm of the turn of the XIX and XX century, as well as potential variants of this norm; 2. phenomena remaining within the norm, but, e.g., typical of the period in question or of a given text. The selection comprised the following elements: phenomena considered then as rare, archaisms, dialectal phrases, localismus and neologismus. They were all described in chapters on phonetics, morphology, syntax and vocabulary. The kind of the elements selected and the extent to which they permeat the texts point out to moderate dialectizing and slight archaizing of poetry. Moderate permeation with innovations is also present. All these devices were, in accordance with the tradition, introduced primarily by means of lexical phenomena. Most of the linguistic means used can be classified as traditional and known in literature. The quality of these means and the way they are used make it possible for Kasprowicz to be counted among poets using interesting language.
Research based on a treebank is active for many natural language applications. However, the work to build a large scale treebank is laborious and tedious. This paper proposes a probabilistic chunker to help the development of a partially bracketed corpus. The chunker partitions the part-of-speech sequence into segments called chunks. Rather than using a treebank as our training corpus, a corpus which is tagged with part-of-speech information only is used. The experimental results show the probabilistic chunker has more than 92% correct rate in outside test. The well-formed partially bracketed corpus is a milestone in the development of a treebank. Besides, the simple but effective chunker can also be applied to many natural language applications.