The ICE-AUS corpus is a 1 m.word corpus of transcribed spoken and written Australian English from 1992-1995. Its internal structure with 500 samples (60% speech, 40%writing) matches that of other ICE
Compiled in October 2019, a Matukar Panau corpus of over 127,000 words produced from newly and previously collected data, including words in context, speaker metadat, file metadata and where available