COCT 口語語料庫2016
Metadata for COCT 口語語料庫2016
Corpus title COCT 口語語料庫2016
CQPweb's short handles for this corpus bl / BL
Total number of texts in corpus 1,180
Total word tokens in all corpus texts 4,383,093
Word types in the corpus 84,869
Standardised type:token ratio (1,000-token basis) Cannot be displayed (STTR not cached)
Non-standardised type:token ratio 0.0194 types per token
Text metadata and word-level annotation
The database stores the following information for each text in the corpus: program
The primary classification of texts is based on: program
Words in this corpus are annotated with: pos
The primary word-level annotation scheme is: No primary word-level annotation scheme has been set