COCT 口語語料庫2017
Metadata for COCT 口語語料庫2017
Corpus title COCT 口語語料庫2017
CQPweb's short handles for this corpus bl2017 / BL2017
Total number of texts in corpus 3,146
Total word tokens in all corpus texts 11,707,405
Word types in the corpus 137,060
Standardised type:token ratio (1,000-token basis) Cannot be displayed (STTR not cached)
Non-standardised type:token ratio 0.0117 types per token
Text metadata and word-level annotation
The database stores the following information for each text in the corpus: There is no text-level metadata for this corpus.
The primary classification of texts is based on: A primary classification scheme for texts has not been set.
Words in this corpus are annotated with: pos
The primary word-level annotation scheme is: No primary word-level annotation scheme has been set