The model for lemmatisation of standard Croatian was built with the [CLASSLA-Stanza tool](https://github.com/clarinsi/classla) by training on the [hr500k training corpus](http://hdl.handle.net/11356/1
(1) The BJCMC Corpus) The Beijing Child Mandarin Corpus (BJCMC) was constructed to address the absence of systematic documentation of child Mandarin speech in naturalistic contexts at preschool age (3