929 resultados para second language processing
Resumo:
Es un estudio que plantea un aspecto traductológico específico: la traducción a la segunda lengua de quien traduce. A partir del análisis de una traducción al español de The Cat in Ihe Hat, de Dr. Seuss, obra de literatura infantil, se señalan ciertas ventaj as metodológicas que tienen que ver con la competencia cultural, la comprensión apropiada del texto original y un acercamiento lingüístico más consciente de la lengua terminal.This study focuses on translation directionality; in particular, translating to one's second language. Based on a Spanish translation of the children's book The Cat in in the Hat, by Dr. Seuss, this article discusses certain methodological advantages of translating to one's second language: cultural competency, an appropriate understanding of the source text, and an enhanced linguistic awareness of the target language.
Resumo:
Tese (doutorado)—Universidade de Brasília, Instituto de Letras, Departamento de Linguística, Português e Línguas Clássicas, Programa de Pós-Graduação em Linguística, 2016.
Resumo:
Ce mémoire tente de répondre à une problématique très importante dans le domaine de recrutement : l’appariement entre offre d’emploi et candidats. Dans notre cas nous disposons de milliers d’offres d’emploi et de millions de profils ramassés sur les sites dédiés et fournis par un industriel spécialisé dans le recrutement. Les offres d’emploi et les profils de candidats sur les réseaux sociaux professionnels sont généralement destinés à des lecteurs humains qui sont les recruteurs et les chercheurs d’emploi. Chercher à effectuer une sélection automatique de profils pour une offre d’emploi se heurte donc à certaines difficultés que nous avons cherché à résoudre dans le présent mémoire. Nous avons utilisé des techniques de traitement automatique de la langue naturelle pour extraire automatiquement les informations pertinentes dans une offre d’emploi afin de construite une requête qui nous permettrait d’interroger notre base de données de profils. Pour valider notre modèle d’extraction de métier, de compétences et de d’expérience, nous avons évalué ces trois différentes tâches séparément en nous basant sur une référence cent offres d’emploi canadiennes que nous avons manuellement annotée. Et pour valider notre outil d’appariement nous avons fait évaluer le résultat de l’appariement de dix offres d’emploi canadiennes par un expert en recrutement.
Resumo:
Prominent views in second language acquisition suggest that the age of L2 learning is inversely correlated with native-like pronunciation (Scovel, 1988; Birdsong, 1999). The relationship has been defined in terms of the Critical Period Hypothesis, whereby various aspects of neural cognition simultaneously occur near the onset of puberty, thus inhibiting L2 phonological acquisition. The current study tests this claim of a chronological decline in pronunciation aptitude through the examination of a key trait of American English – reduced vowels, or “schwas.” Groups of monolingual, early bilingual, and late bilingual participants were directly compared across a variety of environments phonologically conditioned for vowel reduction. Results indicate that late bilinguals have greater degrees of difficulty in producing schwas, as expected. Results further suggest that the degree of differentiation between schwa is larger than previously identified and that these subtle differences may likely be a contributive factor to the perception of a foreign accent in bilingual speakers.
Resumo:
This investigation focused on the treatment of English deictic verbs of motion by Spanish-English bilinguals in Miami. Although English and Spanish share significant overlap of the spatial deixis system, they diverge in important aspects. It is not known how these verbs are processed by bilinguals. Thus, this study examined Spanish-English bilinguals’ interpretation of the verbs come, go, bring, and take in English. Forty-five monolingual English speakers and Spanish-English bilinguals participated. Participants were asked to watch video clips depicting motion events and to judge the acceptability of accompanying narrations spoken by the actors in the videos. Analyses showed that, in general, monolinguals and bilinguals patterned similarly across the deictic verbs come, bring, go and take. However, they did differ in relation to acceptability of word order for verbal objects. Also, bring was highly accepted by all language groups across all goal paths, possibly suggesting an innovation in its use.
Resumo:
Influential bodies of work in language acquisition studies single out heritage bilingualism as a discrete acquisition process within the bilingualism continuum. In regards to the acquisition of WH-/QU- interrogatives containing prepositional phrases (PP), the present study examined whether heritage speakers (HS) of Brazilian Portuguese (BP) produce preposition stranding (P-stranding) constructions in their heritage language, in contrast to monolingual and adult speakers of BP, where prepositions are pied-piped to form the interrogative. Participants were HS of BP born in the USA and in Brazil, monolinguals, and late bilingual adults. The experiment consisted of an elicited production task and a grammaticality judgment task, both carried out in BP and then in English. Results showed that HS born in the USA use P-stranding in QU- interrogatives productively and systematically, in contrast to the other three groups. Moreover, no evidence of protracted acquisition was found in this group. No signs of attrition were detected among bilinguals.
Resumo:
L’objectif de cette étude qualitative est de décrire et de comprendre le processus décisionnel sous-jacent à la rétroaction corrective d’un enseignant de langue seconde à l’oral. Pour ce faire, elle décrit les principaux facteurs qui influencent la décision de procéder à une rétroaction corrective ainsi que ceux qui sous-tendent le choix d’une technique de rétroaction particulière. Trois enseignantes de français langue seconde auprès d’un public d’adultes immigrants au Canada ont participé à cette recherche. Des séquences complètes d’enseignement ont été filmées puis présentées aux participantes qui ont commenté leur pratique. L’entretien de verbalisation s’est effectué sous la forme d’un rappel stimulé et d’une entrevue. Cet entretien constitue les données de cette étude. Les résultats ont révélé que la rétroaction corrective ainsi que le choix de la technique employée étaient influencés par des facteurs relatifs à l’erreur, à l’apprenant, au curriculum, à l’enseignant et aux caractéristiques des techniques. Ils ont également révélé que l’apprenant est au cœur du processus décisionnel rétroactif des enseignants de langue seconde. En effet, les participantes ont affirmé vouloir s’adapter à son fonctionnement cognitif, à son état affectif, à son niveau de langue et à la récurrence de ses erreurs. L’objectif de cette étude est d’enrichir le domaine de la formation initiale et continue des enseignants de L2. Pour cela, des implications pédagogiques ont été envisagées et la recommandation a été faite de porter à la connaissance des enseignants de L2 les résultats des recherches sur l’efficacité des techniques de rétroaction corrective, particulièrement celles qui prennent en compte les caractéristiques des apprenants.
Resumo:
Ce mémoire tente de répondre à une problématique très importante dans le domaine de recrutement : l’appariement entre offre d’emploi et candidats. Dans notre cas nous disposons de milliers d’offres d’emploi et de millions de profils ramassés sur les sites dédiés et fournis par un industriel spécialisé dans le recrutement. Les offres d’emploi et les profils de candidats sur les réseaux sociaux professionnels sont généralement destinés à des lecteurs humains qui sont les recruteurs et les chercheurs d’emploi. Chercher à effectuer une sélection automatique de profils pour une offre d’emploi se heurte donc à certaines difficultés que nous avons cherché à résoudre dans le présent mémoire. Nous avons utilisé des techniques de traitement automatique de la langue naturelle pour extraire automatiquement les informations pertinentes dans une offre d’emploi afin de construite une requête qui nous permettrait d’interroger notre base de données de profils. Pour valider notre modèle d’extraction de métier, de compétences et de d’expérience, nous avons évalué ces trois différentes tâches séparément en nous basant sur une référence cent offres d’emploi canadiennes que nous avons manuellement annotée. Et pour valider notre outil d’appariement nous avons fait évaluer le résultat de l’appariement de dix offres d’emploi canadiennes par un expert en recrutement.
Resumo:
Several definitions exist that offer to identify the boundaries between languages and dialects, yet these distinctions are inconsistent and are often as political as they are linguistic (Chambers & Trudgill, 1998). A different perspective is offered in this thesis, by investigating how closely related linguistic varieties are represented in the brain and whether they engender similar cognitive effects as is often reported for bilingual speakers of recognised independent languages, based on the principles of Green’s (1998) model of bilingual language control. Study 1 investigated whether bidialectal speakers exhibit similar benefits in non-linguistic inhibitory control as a result of the maintenance and use of two dialects, as has been proposed for bilinguals who regularly employ inhibitory control mechanisms, in order to suppress one language while speaking the other. The results revealed virtually identical performance across all monolingual, bidialectal and bilingual participant groups, thereby not just failing to find a cognitive control advantage in bidialectal speakers over monodialectals/monolinguals, but also in bilinguals; adding to a growing body of evidence which challenges this bilingual advantage in non-linguistic inhibitory control. Study 2 investigated the cognitive representation of dialects using an adaptation of a Language Switching Paradigm to determine if the effort required to switch between dialects is similar to the effort required to switch between languages. The results closely replicated what is typically shown for bilinguals: Bidialectal speakers exhibited a symmetrical switch cost like balanced bilinguals while monodialectal speakers, who were taught to use the dialect words before the experiment, showed the asymmetrical switch cost typically displayed by second language learners. These findings augment Green’s (1998) model by suggesting that words from different dialects are also tagged in the mental lexicon, just like words from different languages, and as a consequence, it takes cognitive effort to switch between these mental settings. Study 3 explored an additional explanation for language switching costs by investigating whether changes in articulatory settings when switching between different linguistic varieties could - at least in part – be responsible for these previously reported switching costs. Using a paradigm which required participants to switch between using different articulatory settings, e.g. glottal stops/aspirated /t/ and whispers/normal phonation, the results also demonstrated the presence of switch costs, suggesting that switching between linguistic varieties has a motor task-switching component which is independent of representations in the mental lexicon. Finally, Study 4 investigated how much exposure is needed to be able to distinguish between different varieties using two novel language categorisation tasks which compared German vs Russian cognates, and Standard Scottish English vs Dundonian Scots cognates. The results showed that even a small amount of exposure (i.e. a couple of days’ worth) is required to enable listeners to distinguish between different languages, dialects or accents based on general phonetic and phonological characteristics, suggesting that the general sound template of a language variety can be represented before exact lexical representations have been formed. Overall, these results show that bidialectal use of typologically closely related linguistic varieties employs similar cognitive mechanisms as bilingual language use. This thesis is the first to explore the cognitive representations and mechanisms that underpin the use of typologically closely related varieties. It offers a few novel insights and serves as the starting point for a research agenda that can yield a more fine-grained understanding of the cognitive mechanisms that may operate when speakers use closely related varieties. In doing so, it urges caution when making assumptions about differences in the mechanisms used by individuals commonly categorised as monolinguals, to avoid potentially confounding any comparisons made with bilinguals.
Resumo:
Neuroimaging research involves analyses of huge amounts of biological data that might or might not be related with cognition. This relationship is usually approached using univariate methods, and, therefore, correction methods are mandatory for reducing false positives. Nevertheless, the probability of false negatives is also increased. Multivariate frameworks have been proposed for helping to alleviate this balance. Here we apply multivariate distance matrix regression for the simultaneous analysis of biological and cognitive data, namely, structural connections among 82 brain regions and several latent factors estimating cognitive performance. We tested whether cognitive differences predict distances among individuals regarding their connectivity pattern. Beginning with 3,321 connections among regions, the 36 edges better predicted by the individuals' cognitive scores were selected. Cognitive scores were related to connectivity distances in both the full (3,321) and reduced (36) connectivity patterns. The selected edges connect regions distributed across the entire brain and the network defined by these edges supports high-order cognitive processes such as (a) (fluid) executive control, (b) (crystallized) recognition, learning, and language processing, and (c) visuospatial processing. This multivariate study suggests that one widespread, but limited number, of regions in the human brain, supports high-level cognitive ability differences. Hum Brain Mapp, 2016. © 2016 Wiley Periodicals, Inc.
Resumo:
L’objectif de cette étude qualitative est de décrire et de comprendre le processus décisionnel sous-jacent à la rétroaction corrective d’un enseignant de langue seconde à l’oral. Pour ce faire, elle décrit les principaux facteurs qui influencent la décision de procéder à une rétroaction corrective ainsi que ceux qui sous-tendent le choix d’une technique de rétroaction particulière. Trois enseignantes de français langue seconde auprès d’un public d’adultes immigrants au Canada ont participé à cette recherche. Des séquences complètes d’enseignement ont été filmées puis présentées aux participantes qui ont commenté leur pratique. L’entretien de verbalisation s’est effectué sous la forme d’un rappel stimulé et d’une entrevue. Cet entretien constitue les données de cette étude. Les résultats ont révélé que la rétroaction corrective ainsi que le choix de la technique employée étaient influencés par des facteurs relatifs à l’erreur, à l’apprenant, au curriculum, à l’enseignant et aux caractéristiques des techniques. Ils ont également révélé que l’apprenant est au cœur du processus décisionnel rétroactif des enseignants de langue seconde. En effet, les participantes ont affirmé vouloir s’adapter à son fonctionnement cognitif, à son état affectif, à son niveau de langue et à la récurrence de ses erreurs. L’objectif de cette étude est d’enrichir le domaine de la formation initiale et continue des enseignants de L2. Pour cela, des implications pédagogiques ont été envisagées et la recommandation a été faite de porter à la connaissance des enseignants de L2 les résultats des recherches sur l’efficacité des techniques de rétroaction corrective, particulièrement celles qui prennent en compte les caractéristiques des apprenants.
Resumo:
O processamento de linguagem natural e as ontologias são ferramentas cuja interação permite uma melhor compreensão dos dados armazenados. Este trabalho, ao associar estas duas áreas aos elementos disponíveis numa base de dados prosopográfica, tornou possível identificar e classificar relacionamentos entre setores de ocupação na forma como eram designados na época, setores de atividade num formato mais próximo do de hoje e o estatuto social que essas incumbências tinham na sociedade coeva. Os dados utilizados são sobretudo de membros do Santo Ofício – do século XVI ao século XVIII. Para atingir este objetivo utilizaram-se algumas descrições textuais de ocorrências da época e outras pouco estruturadas, disponíveis no repositório SPARES. A aplicação de processamento de linguagem natural (remoção de stopwords e aplicação de stemming), conjugada com a construção de duas ontologias, tornou possível classificar esses dados, permitindo consultas mais eficazes. Ao contribuir para a classificação automática de dados históricos, propõem-se metodologias que podem ser aplicadas em dados de qualquer outra área do conhecimento, especialmente as que lidam com as variáveis de tempo e espaço de forma mais intensa; Abstract: OntoSPARES: from natural language to ontologies Contributions to the automatic classification of historical data (16th-18th centuries) The interaction between the natural language processing and ontologies are tools allowing a better understanding of the data stored. This work, by combining these two areas to the elements available in a prosopographic database, has made possible to identify and classify relationships between occupations of many individuals (in general Holy Office members of the 16th-18th centuries). To achieve this goal the data used was gathered in SPARES repository, including some textual descriptions of the time occurrences. They are all few structured. The application of natural language processing (stopwords removal and stemming application), combined with the construction of two ontologies, made possible to classify those data, allowing a more effective search. By contributing to the automatic classification of historical data, this thesis proposes methodologies that can be applied to data from any other field of knowledge, specially data dealing with time and space variables.
Resumo:
A evolução tecnológica tem provocado uma evolução na medicina, através de sistemas computacionais voltados para o armazenamento, captura e disponibilização de informações médicas. Os relatórios médicos são, na maior parte das vezes, guardados num texto livre não estruturado e escritos com vocabulário proprietário, podendo ocasionar falhas de interpretação. Através das linguagens da Web Semântica, é possível utilizar antologias como modo de estruturar e padronizar a informação dos relatórios médicos, adicionando¬ lhe anotações semânticas. A informação contida nos relatórios pode desta forma ser publicada na Web, permitindo às máquinas o processamento automático da informação. No entanto, o processo de criação de antologias é bastante complexo, pois existe o problema de criar uma ontologia que não cubra todo o domínio pretendido. Este trabalho incide na criação de uma ontologia e respectiva povoação, através de técnicas de PLN e Aprendizagem Automática que permitem extrair a informação dos relatórios médicos. Foi desenvolvida uma aplicação, que permite ao utilizador converter relatórios do formato digital para o formato OWL. ABSTRACT: Technological evolution has caused a medicine evolution through computer systems which allow storage, gathering and availability of medical information. Medical reports are, most of the times, stored in a non-structured free text and written in a personal way so that misunderstandings may occur. Through Semantic Web languages, it’s possible to use ontology as a way to structure and standardize medical reports information by adding semantic notes. The information in those reports can, by these means, be displayed on the web, allowing machines automatic information processing. However, the process of creating ontology is very complex, as there is a risk creating of an ontology that not covering the whole desired domain. This work is about creation of an ontology and its population through NLP and Machine Learning techniques to extract information from medical reports. An application was developed which allows the user to convert reports from digital for¬ mat to OWL format.
Resumo:
Question Answering systems that resort to the Semantic Web as a knowledge base can go well beyond the usual matching words in documents and, preferably, find a precise answer, without requiring user help to interpret the documents returned. In this paper, the authors introduce a Dialogue Manager that, through the analysis of the question and the type of expected answer, provides accurate answers to the questions posed in Natural Language. The Dialogue Manager not only represents the semantics of the questions, but also represents the structure of the discourse, including the user intentions and the questions context, adding the ability to deal with multiple answers and providing justified answers. The authors’ system performance is evaluated by comparing with similar question answering systems. Although the test suite is slight dimension, the results obtained are very promising.
Resumo:
Bangla OCR (Optical Character Recognition) is a long deserving software for Bengali community all over the world. Numerous e efforts suggest that due to the inherent complex nature of Bangla alphabet and its word formation process development of high fidelity OCR producing a reasonably acceptable output still remains a challenge. One possible way of improvement is by using post processing of OCR’s output; algorithms such as Edit Distance and the use of n-grams statistical information have been used to rectify misspelled words in language processing. This work presents the first known approach to use these algorithms to replace misrecognized words produced by Bangla OCR. The assessment is made on a set of fifty documents written in Bangla script and uses a dictionary of 541,167 words. The proposed correction model can correct several words lowering the recognition error rate by 2.87% and 3.18% for the character based n- gram and edit distance algorithms respectively. The developed system suggests a list of 5 (five) alternatives for a misspelled word. It is found that in 33.82% cases, the correct word is the topmost suggestion of 5 words list for n-gram algorithm while using Edit distance algorithm the first word in the suggestion properly matches 36.31% of the cases. This work will ignite rooms of thoughts for possible improvements in character recognition endeavour.