Understanding Indonesian: developing a machine-usable grammar, dictionary and corpus. Australia's relationship with Indonesia is of great significance. The need for good relationships founded on appreciation of the range of societies and views in modern Indonesia is widely acknowledged. A better knowledge of the languages is essential for this, and so are fast, efficient information gathering systems for processing multilingual sources (including Indonesian text), that can analyse large volumes ....Understanding Indonesian: developing a machine-usable grammar, dictionary and corpus. Australia's relationship with Indonesia is of great significance. The need for good relationships founded on appreciation of the range of societies and views in modern Indonesia is widely acknowledged. A better knowledge of the languages is essential for this, and so are fast, efficient information gathering systems for processing multilingual sources (including Indonesian text), that can analyse large volumes of text. The skills to build such systems exist internationally. Through collaboration with established international teams, we plan to transfer cutting-edge skills in the development of machine-useable grammars to Australian researchers, and to create the language resources essential for understanding Indonesian.Read moreRead less
Linkage Infrastructure, Equipment And Facilities - Grant ID: LE100100211
Funder
Australian Research Council
Funding Amount
$650,000.00
Summary
The Big Australian Speech Corpus: An audio-visual speech corpus of Australian English. Contemporary speech science and technology are driven by the availability of large speech corpora. While audio databases exist for languages spoken in America, Europe and Japan, there is currently no large auditory-visual database of spoken language, and certainly not one for Australian English. Here we will establish the Big Australian Speech Corpus, which will support a speech science research and developmen ....The Big Australian Speech Corpus: An audio-visual speech corpus of Australian English. Contemporary speech science and technology are driven by the availability of large speech corpora. While audio databases exist for languages spoken in America, Europe and Japan, there is currently no large auditory-visual database of spoken language, and certainly not one for Australian English. Here we will establish the Big Australian Speech Corpus, which will support a speech science research and development using Australian English and facilitate the development of Australian speech technology applications from automatic speech recognition and text-to-speech synthesis used in taxi and other ordering services, to hearing prostheses and talking head aids for learning-impaired children, and a range of security and forensic applications.Read moreRead less