Corpus Search Tutorial: English-Corpora.org Queries

Added:

Search Basics
Using POS Tags
Refined Searches
Decade Analysis
Lemma Search
Collocation Queries
Phrase Patterns
Exclusion Filters
Sequence Analysis

Search Basics

0:00
Playing Section
  • 1

    Learn to navigate the English-corpora.org interface and log in.

  • 2

    Use the COCA corpus to find the frequency of specific words like 'rad'.

Understanding the basic definition of a 'corpus' and how corpus linguistics differs from traditional, prescriptive grammar.
Familiarity with English parts of speech (POS) categories and grammatical concepts (e.g., lexical vs. functional words).
Basic conceptual understanding of database querying, search wildcards (like asterisks), and boolean logic.
Analyzing collocates and understanding association metrics like Mutual Information (MI) scores and t-scores.
Conducting register and diachronic (historical) analyses to compare language use across different genres and time periods.
Interpreting concordance lines (Key Word in Context) to analyze semantic prosody and syntactic patterns.
Applying corpus-derived data to practical domains such as lexicography, language pedagogy, or sentiment analysis.
683 views21likes21:12@ekbphd3200Original Release: 2024-09-09

This video tutorial demonstrates fundamental search techniques for English corpora (COCA, COHA, NOW) on English-corpora.org, including searching for word frequencies by genre/register, using part-of-speech tags (POS) with uppercase for full words and lowercase for specific word forms, searching for lemmas (base word forms) using capital letters, finding words following or preceding other words using asterisks, and using the pipe operator for multiple conditions. The instructor shows how to switch between corpora, interpret frequency charts, and analyze linguistic patterns across different time periods and genres.