Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Online Corpora

View through CrossRef
Abstract An online corpus is a structured digital collection of texts accessible via the Internet, widely used in linguistic research and language studies. Online corpora can be general, specialized, or parallel, supporting areas such as discourse analysis, sociolinguistics, language teaching and learning, and lexicography. They differ from offline corpora, which are stored locally and provide greater control over data but lack the accessibility and frequent updates of online resources. Some corpora, such as COCA and NOW, are monitor corpora, regularly updated to capture ongoing changes in language use. Online corpora also vary in design: some are primarily research‐oriented, offering advanced search and analysis tools (e.g., Sketch Engine, AntConc, #LancsBox), while others are platforms with a pedagogical focus, providing user‐friendly access to corpora and tools for students and teachers (e.g., SkELL, CorpusMate). Their advantages include accessibility, scope, diversity, and integrated analytical tools; limitations include dependence on Internet connectivity and reduced control over content. Online corpora have significantly influenced both theoretical and applied linguistics, contributing to dictionary design, the development of teaching materials, and feedback in language learning. This entry introduces the main types, uses, advantages, and limitations of online corpora, situating them within the broader field of corpus linguistics.
Title: Online Corpora
Description:
Abstract An online corpus is a structured digital collection of texts accessible via the Internet, widely used in linguistic research and language studies.
Online corpora can be general, specialized, or parallel, supporting areas such as discourse analysis, sociolinguistics, language teaching and learning, and lexicography.
They differ from offline corpora, which are stored locally and provide greater control over data but lack the accessibility and frequent updates of online resources.
Some corpora, such as COCA and NOW, are monitor corpora, regularly updated to capture ongoing changes in language use.
Online corpora also vary in design: some are primarily research‐oriented, offering advanced search and analysis tools (e.
g.
, Sketch Engine, AntConc, #LancsBox), while others are platforms with a pedagogical focus, providing user‐friendly access to corpora and tools for students and teachers (e.
g.
, SkELL, CorpusMate).
Their advantages include accessibility, scope, diversity, and integrated analytical tools; limitations include dependence on Internet connectivity and reduced control over content.
Online corpora have significantly influenced both theoretical and applied linguistics, contributing to dictionary design, the development of teaching materials, and feedback in language learning.
This entry introduces the main types, uses, advantages, and limitations of online corpora, situating them within the broader field of corpus linguistics.

Related Results

UMA PROPOSTA DE WORKFLOW PARA CONSTRUÇÃO DE CORPUS DIGITAL EM LÍNGUA DE SINAIS
UMA PROPOSTA DE WORKFLOW PARA CONSTRUÇÃO DE CORPUS DIGITAL EM LÍNGUA DE SINAIS
Os corpora de línguas de sinais disponíveis atualmente em pesquisas linguísticas e em sites para acesso livre são constituídos por um módulo de gravação feita em vídeo, pois os dad...
A Taste for Corpora
A Taste for Corpora
The eleven contributions to this volume, written by expert corpus linguists, tackle corpora from a wide range of perspectives and aim to shed light on the numerous linguistic and p...
A Research Overview of Corpus-Assisted Enhancement of English Writing Proficiency
A Research Overview of Corpus-Assisted Enhancement of English Writing Proficiency
With the deep development of big data and artificial intelligence technologies, corpus research has garnered increasing attention and recognition. Initially, corpora were collectio...
Comparable Corpora: Compilation Methods and Areas of Application
Comparable Corpora: Compilation Methods and Areas of Application
Comparable corpora and their application in research have been an object of interest since the 1990s. Following the establishment of the annual workshop series “Building and Using ...
Žanrovska analiza pomorskopravnih tekstova i ostvarenje prijevodnih univerzalija u njihovim prijevodima s engleskoga jezika
Žanrovska analiza pomorskopravnih tekstova i ostvarenje prijevodnih univerzalija u njihovim prijevodima s engleskoga jezika
Genre implies formal and stylistic conventions of a particular text type, which inevitably affects the translation process. This „force of genre bias“ (Prieto Ramos, 2014) has been...
Phonetic Corpora
Phonetic Corpora
Abstract Technological advancements in recording, storage, and processing have reshaped the study of spoken language data over the past decades. Phonetic corpora ...
NEARSIDE: Structured kNowledge Extraction frAmework from SpecIes DEscriptions
NEARSIDE: Structured kNowledge Extraction frAmework from SpecIes DEscriptions
Species descriptions are stored in textual form in corpora such as in floras and faunas, but this large amount of information cannot be used directly by algorithms, nor can it be l...
Spoken Corpora of Slavic Languages
Spoken Corpora of Slavic Languages
AbstractSpoken corpora are collections of transcribed and annotated audio and /or video recordings of languages or language varieties. The aim of this paper is to present an overvi...

Back to Top