Monitor Corpora. PressCoronaVírus
Identification
- Project identification: Monitor Corpora. PressCoronaVírus
- Coordination: Raquel Silva and Margarida Ramos (Lexicology, Lexicography and Terminology research group)
- Start date: August 2021
- Funding: Verão com Ciência, FCT – Portuguese Science and Technology National Foundation
Description
The Monitor Corpora Project falls within the scope of the COVID-19 Collaborative Glossary, a resource consisting of the terminology used by official Health bodies, professionals and scientists from the sector, as well as the media and social networks.
The aim of Monitor Corpora Project is the monitoring of a web corpus, specifically composed of texts from the media, collected in the context of the pandemic situation created by the Coronavirus – the PressCoronaVirus. This corpus gathers texts produced in European Portuguese for consultation, collection and processing of linguistic data. The texts were extracted from reliable journalistic sources whose textual productions are updated, which is of utmost importance for the current studies on neology in contemporary Portuguese language. Finally, the corpus is open to the inclusion of new texts in order to enrich the corpus of analysis.
The PressCoronaVirus comprises texts published between November 2020 and July 2021, and was processed with SketchEngine with the purpose of identifying neological lexical constructions associated with non-specialized communication situations related to the emergence of the pandemic. It is thus possible to observe in real time (synchrony) and within a certain time span (diachrony), the lexical and neological productivity resulting from the social, cultural and economic context caused by COVID-19. We intend that the outcome obtained with the methodology will allow us to: (i) keep the corpus updated; (ii) feed the COVID-19 collaborative glossary; (iii) and make the corpus available in open access. The corpus files and respective documentation (metadata files) will be made available under the Attribution-NonCommercial-NoDerivatives 4.0 International license of the Creative Commons.
Team
Raquel Silva
Margarida Ramos
Sandro Loupa
Menu < back
- Projects
- Ongoing projects
- HEREDITARY – HetERogeneous sEmantic Data Integration for guT-brAin interplay
- TTC-CPLP – Terminologias Técnicas e Científicas para a CPLP
- CHAMUÇA – Portuguese and South Asian Lexicon Archive
- e-Term ANCV – Recurso terminológico jurídico-parlamentar digital Assembleia Nacional de Cabo Verde
- NObarriers2Health: Reducing language and cultural barriers through machine translation literacy for inclusive multilingual health communication
- EPISTRAN – Epistemic Translation: Towards an Ecology of Knowledges
- DiTo – Didática do Texto
- REDGRAM – Digital Resources for Education – Grammatical Pathways
- iRead4Skills – Intelligent Reading Improvement System for Fundamental and Transversal Skills Development
- Active Citizenship Through Dialogue in Virtual teacher communities
- ProPerL2 – Production and Perception in L2 speech learning
- Heritage Languages go to School: The interplay of (extra)linguistic factors in successful language development
- Investigating the impact of implicit and explicit instruction on phonological acquisition in a second language
- LAUA – Language Attrition and Ultimate Attainment
- CORRELATE – Corpora and Lexical and Terminological Resources
- ANACOREX – Anafora y expresiones referenciales en el bilinguismo: triangulando enfoques de corpus y experimentales
- Caring Communication: gene therapy in the context of hemophilia
- CoRaLHis – Comparing Romance Languages through History: building a multilingual parallel diachronic corpus (13th-18th C.)
- MorDigital – Digitisation of Diccionario da Lingua Portugueza by António de Morais Silva
- Western Sephardic Diaspora Roadmap
- EXPRIMI
- Language and literacy at school – the contribution of metasyntactic abilities to reading comprehension development
- G&T.Comenta
- COVID-19 Collaborative Glossary
- TERMVEST – The Clothing Terminology: European Portuguese version
- Digital Edition of the “Vocabulário Ortográfico da Língua Portuguesa” (VOLP-1940)
- PIPALE – Preventive Intervention Project for Learning to Read and Write
- POR Nível – Design and validation of a placement test to PFL
- ANACOR: A corpus-based approach to anaphora resolution in second language acquisition: beyond the interfaces
- Cultural Heritage Lexicon
- Concluded projects
- ELEXIS – European Lexicographic Infrastructure
- Humanities Going Digital (HUGOD)
- LL2DS – Linking Linguistics to Data Science
- QuILL – Quality in Language Learning
- Corpus Linguístico & Avatar para a Língua Gestual Portuguesa
- Monitor Corpora. PressCoronaVírus
- Com@Rehab – Communication for interactive rehabilitation in virtual reality
- Read4Succeed: Improving migrant, refugee and from deprived neighbourhood children reading skills through an Animal Assisted Reading program
- Project GiroFLE
- OrthoDef
- European Portuguese-Standard Arab Dictionary
- MOCOLANG-O – MOdélisation COnceptuelle des troubles (du LANGage et de la communication) en Orthophonie
- Romance clitics in diachrony. An integrated approach
- Portuguese Literature Corpus for Distant Reading
- ALPROF – Automatic Assessment of Language Proficiency for Migrant Integration
- CLARIN CLUNL
- Utopia, Food and the Future
- Development of syntactic structures in Portuguese and French monolingual and bilingual acquisition
- The Case of Grammatical Relations
- BlackBox – a Collaborative Platform to Document Performance Composition: from conceptual structures in the backstage to customizable visualizations in the front-end
- Promotion of scientific literacy
- PerGRam – Percursos para o ensino da gramática nos primeiros anos de escolaridade
- Knowledge Organisation Proposal within the scope of infertility: the role of Terminology
- Subordination in Medieval Portuguese
- Crosslinguistic and Crosspopulation approaches to the Acquisition of Dependencies
- Syntactic and lexical factors in processing complexity
- SIERA – Integrating Sina Institute into the European Research Area
- Syntactic Dependencies from 3 to 10
- Events and subevents in Capeverdean
- TKB – Transmedia Knowledge Base for Contemporary Dance
- Research network projects
- ELEXIS Association
- PhraConRep – A Multilingual Repository of Phraseme Constructions in Central and Eastern European Languages
- Y-JustLang – Justice to youth language needs
- ENEOLI – European Network On Lexical Innovation
- Consortium Huma-Num ARIANE
- GRAFE’Maire
- UniDive – Universality, diversity and idiosyncrasy in language technology
- Metalex – International Metalexicography Network
- @ Cientista Regressa à Escola
- CLIL in Languages Other Than English
- NexusLinguarum – European network for Web-centred linguistic data science
- Distant Reading for European Literary History
- HL2C – Heritage Language Consortium
- KEYSTONE – Semantic Keyword-Based Search on Structures Data Sources
- ARLE – International Association for Research in L1 Education
- ENeL – European Network of e-Lexicography
- GraMaLL – Grasping Meaning Across Languages and Learners
- Language Impairment in a Multilingual Society: Linguistic Patterns and the Road to Assessment
- GIRTraduvino – Grupo de Investigación Reconocido sobre la Lengua de la Vid Y el Vino y su Traducción
- Value for Health CoLAB
- Infrastructures
- Services provision