COCT 書面語語料庫2020
Metadata for COCT 書面語語料庫2020
Corpus title COCT 書面語語料庫2020
CQPweb's short handles for this corpus yl2020_v3 / YL2020_V3
Total number of texts in corpus 114,606
Total word tokens in all corpus texts 266,486,260
Word types in the corpus 2,773,341
Standardised type:token ratio (1,000-token basis) 0.4317 types per token
Non-standardised type:token ratio 0.0104 types per token
Text metadata and word-level annotation
The database stores the following information for each text in the corpus: There is no text-level metadata for this corpus.
The primary classification of texts is based on: A primary classification scheme for texts has not been set.
Words in this corpus are annotated with: pos
The primary word-level annotation scheme is: No primary word-level annotation scheme has been set