| Metadata for COCT 口語語語料庫2021 |
| Corpus title |
COCT 口語語語料庫2021 |
| CQPweb's short handles for this corpus |
da2021_all_v1 / DA2021_ALL_V1 |
| Total number of texts in corpus |
7,021 |
| Total word tokens in all corpus texts |
24,303,812 |
| Word types in the corpus |
227,839 |
| Standardised type:token ratio (1,000-token basis) |
Cannot be displayed (STTR not cached) |
| Non-standardised type:token ratio |
0.0094 types per token |
| Text metadata and word-level annotation |
| The database stores the following information for each text in the corpus: |
There is no text-level metadata for this corpus. |
| The primary classification of texts is based on: |
A primary classification scheme for texts has not been set. |
| Words in this corpus are annotated with: |
pos |
| The primary word-level annotation scheme is: |
No primary word-level annotation scheme has been set |