article Open access

When Data Meet Tools: Using the Monitor Corpus for the Analysis of Language Development

  • Journal of Linguistics/Jazykovedný casopis
  • De Gruyter Open
Research footprint

At a glance

Citations
1
References
4
Comments
0
Paper overview

Öz

Abstract The aim of this paper is to introduce an infrastructure developed within the HiČKoK project to enable full-fledged corpus-based diachronic research of Czech. The individual sections of the paper present the components of this infrastructure, which links well-balanced, representative and annotated data with tailor-made tools for diachronic research. The forthcoming monitor corpus, covering the entire period of written Czech, along with its composition and annotation strategies, is briefly introduced. In the following sections, the potential of the application and its four modules—simple query, comparison, time-based associations, and diachronic collocations—are demonstrated through mini case studies. Combining large-scale data (as representative as possible) with a tool that enhances standard corpus functionalities, enriches them with a diachronic perspective, and enables result visualization makes diachronic research on language change more accessible and comprehensive.

Record transparency

Publication details

DOI
10.2478/jazcas-2025-0014
OpenAlex
W7108340090
Document type
article
Language
EN
Source
Journal of Linguistics/Jazykovedný casopis
Last metadata update
Community

Comments

Oturum Açın to join the discussion.

  1. No comments yet. Start the discussion.