Search results
Results from the WOW.Com Content Network
Corpus linguistics is an empirical method for the study of language by way of a text corpus (plural corpora). [1] Corpora are balanced, often stratified collections of authentic, "real world", text of speech or writing that aim to represent a given linguistic variety . [ 1 ]
Corpus-assisted discourse studies (abbr.: CADS) is related historically and methodologically to the discipline of corpus linguistics.The principal endeavor of corpus-assisted discourse studies is the investigation, and comparison of features of particular discourse types, integrating into the analysis the techniques and tools developed within corpus linguistics.
The economy principle in linguistics, also known as linguistic economy, is a functional explanation of linguistic form. It suggests that the organization of phonology , morphology , lexicon and syntax is fundamentally based on a compromise between simplicity and clarity, two desirable but to some extent incompatible qualities.
In order to make the corpora more useful for doing linguistic research, they are often subjected to a process known as annotation. An example of annotating a corpus is part-of-speech tagging, or POS-tagging, in which information about each word's part of speech (verb, noun, adjective, etc.) is added to the corpus in the form of tags.
The economics of language is an emerging field of study concerning a range of topics such as the effect of language skills on income and trade, the costs and benefits of language planning options, the preservation of minority languages, etc. [1] [2] It is relevant to analysis of language policy.
Text corpora (singular: text corpus) are large and structured sets of texts, which have been systematically collected.Text corpora are used by corpus linguists and within other branches of linguistics for statistical analysis, hypothesis testing, finding patterns of language use, investigating language change and variation, and teaching language proficiency.
Corpus languages are studied using the methods of corpus linguistics, but corpus linguistics can also be used (and is commonly used) for the study of the writings and other records of living languages. Not all extinct languages are corpus languages, since there are many extinct languages in which few or no writings or other records survive.
TaLC (Teaching and Language Corpora) is a biennial conference that is a platform for corpus-based research that has a pedagogical focus. CorpusCALL is a special interest group within EuroCALL and is mostly active through its Facebook group. The online teaching journal, Humanising Language Teaching hosts a section called Corpus Ideas.