Corpus-based research on AusE has been facilitated, over the past half century or so, by the compilation of purpose-built, readily accessible Australian corpora. The first significant corpora of AusE to appear, both comprising around one million words, were the Australian Corpus of English (ACE), designed as a parallel to the British and American ‘Brown family’ corpora, and the Australian component of the International Corpus of English (ICE-AUS), where inclusion of spoken texts expanded the possibilities for exploring register and inter-varietal variation (e.g. Collins 2009).
In addition to these synchronic corpora, several diachronic corpora have been compiled and used, including the Corpus of Oz Early English (COOEE), comprising texts of various kinds from the period 1788-1900 (see Fritz 2007), and ‘AusBrown’, covering the period from 1931 to 2006 (e.g. Collins & Yao 2020). A more specialised corpus focusing on parliamentary language is the Australian Diachronic Hansard Corpus (ADHC), sampled across the period from 2001 to 2015 (e.g. Kruger & Smith 2018). A number of the aforementioned corpora were brought together to form the Australian National Corpus collection (Musgrave & Haugh 2020), and this suite of resources is now curated by the Language Data Commons of Australia (LDaCA). LDaCA’s holdings include a number of more specialised corpora (e.g. the La Trobe Corpus of Spoken AusE, and corpora of Auslan).
This themed session will build on and extend the existing tradition of synchronic and diachronic corpus-based research on Australian English (AusE). We welcome papers that fill gaps in the corpus coverage of AusE, including those that address such topics as the current state of play and future directions and/or present information on corpus-related projects, and those that use data derived from established and/or newly created corpora to examine linguistic aspects of AusE (syntactic, morphological, lexical, phonological, pragmatic or historical) or sociolinguistic aspects (considering differences of gender, age, socioeconomic status, ethnicity, or region as well as social networks). Contributions may explore aspects of variation and change either within AusE or in comparison with other World Englishes. Papers are also welcome on studies that make use of less conventional corpus-like ‘fuzzy sets’ of data, such as survey data or curated collections of historical correspondence. The session will incorporate a moderated panel discussion about changing corpus materials and methods.
We anticipate that this themed session will provide insights into the progression of AusE towards linguistic independence as a World English, and how it is diversifying under the forces of ever-shifting sociodemographics and modes of communication. The research presented will canvas the range and potential of existing AusE corpora, and give direction and inspiration to further research and corpus-building projects.