| Metadata for COCT 書面語語料庫2020 |
| Corpus title |
COCT 書面語語料庫2020 |
| CQPweb's short handles for this corpus |
yl2020_v3 / YL2020_V3 |
| Total number of texts in corpus |
114,606 |
| Total word tokens in all corpus texts |
266,486,260 |
| Word types in the corpus |
2,773,341 |
| Standardised type:token ratio (1,000-token basis) |
0.4317 types per token |
| Non-standardised type:token ratio |
0.0104 types per token |
| Text metadata and word-level annotation |
| The database stores the following information for each text in the corpus: |
There is no text-level metadata for this corpus. |
| The primary classification of texts is based on: |
A primary classification scheme for texts has not been set. |
| Words in this corpus are annotated with: |
pos |
| The primary word-level annotation scheme is: |
No primary word-level annotation scheme has been set |