Determinants of Audio-Visual effects in degraded and non-degraded speech. Seeing a speaker's face can affect the perception of their speech in a number of ways. This project proposes a detailed comparison of factors that affect Audio-Visual (AV) facilitation of degraded speech detection and identification. Detection-based tasks should be more sensitive to signal based correlations whereas identification-based effects more sensitive to complementary information. The significance of the current pr ....Determinants of Audio-Visual effects in degraded and non-degraded speech. Seeing a speaker's face can affect the perception of their speech in a number of ways. This project proposes a detailed comparison of factors that affect Audio-Visual (AV) facilitation of degraded speech detection and identification. Detection-based tasks should be more sensitive to signal based correlations whereas identification-based effects more sensitive to complementary information. The significance of the current proposal is that it offers both a strategy and a connected series of experiments for determining key behavioural constraints on AV speech integration. Understanding AV interactions will build links between neurophysiological processes and coherent perception and have important implications for AV application.Read moreRead less
The Extinction Of Conditioned Fear And Its Implications For Cue Exposure Therapy
Funder
National Health and Medical Research Council
Funding Amount
$322,430.00
Summary
This project studies extinction of Pavlovian conditioned fear reactions in rats. Extinction of these reactions is an animal model for exposure therapy used in the treatment of anxiety disorders in people. In exposure therapy, the patient, aided by the clinician, confronts trauma-related cues in the absence of any overt danger. The intention of this therapy is to reduce the ability of the trauma-related cues to provoke the fear reactions that are undermining the patient's quality of life. In Pavl ....This project studies extinction of Pavlovian conditioned fear reactions in rats. Extinction of these reactions is an animal model for exposure therapy used in the treatment of anxiety disorders in people. In exposure therapy, the patient, aided by the clinician, confronts trauma-related cues in the absence of any overt danger. The intention of this therapy is to reduce the ability of the trauma-related cues to provoke the fear reactions that are undermining the patient's quality of life. In Pavlovian conditioning, subjects (typically rats) are exposed to a signaling relation between an initially neutral stimulus (e.g., a noise) and a feared outcome (e.g., foot shock). When later repeatedly exposed to the initially neutral but now feared stimulus (the noise) in the absence of the feared outcome, the fear reactions it acquired progressively decline until eventually it fails to elicit any such reactions. The fear reactions are said to have been extinguished. There has been significant progress in understanding the psychological processes and neural mechanisms underlying the acquisition of fear reactions, but much less is known about the processes and mechanisms underlying the extinction of these reactions. The project has two general objectives. The first is to determine the conditions of extinction training that promote long-term loss of fear reactions. The second objective is to determine how the brain controls this extinction of learned fear. Achieving these aims will be significant for two reasons. First, it will contribute to understanding the mechanisms by which animals (including people) learn to adjust their behaviour to bring it into line with the current relations that exist between events in the world. Second, it will provide important information about how such adjustment is facilitated or impaired across extinction training and, thereby, contribute towards understanding both the successes and failures of cue exposure therapy for fear-related disorders.Read moreRead less
Robust speech recognition in realistic hostile environments. Australia leads the world in the adoption of speech recognition technology but sadly lags in the development of the fundamental advances in the area. This research will help propel Australia to the forefront of new innovations in speech recognition technology and contributions to fundamental science. Our project will provide an excellent training ground for graduate students and researchers, with the real possibility of significant com ....Robust speech recognition in realistic hostile environments. Australia leads the world in the adoption of speech recognition technology but sadly lags in the development of the fundamental advances in the area. This research will help propel Australia to the forefront of new innovations in speech recognition technology and contributions to fundamental science. Our project will provide an excellent training ground for graduate students and researchers, with the real possibility of significant commercial benefit to the nation. The deployment of our system in the community will greatly enhance the defence and police forces ability for surveillance and security, and will provide new assistive aids to improve the quality of life and safety for the elderly and disabled.Read moreRead less
Enhanced Multilingual Speaker Recognition through the Incorporation of High-Level Features, Late Fusion and Discriminative Classification Methods. The development of robust multilingual speaker recognition systems will benefit the community through the elimination of fraud incurred by financial institutions and customers by enabling several person authentication applications such as: voice based signatures and document issuance; credit card verification by voice and secure over-the-phone financi ....Enhanced Multilingual Speaker Recognition through the Incorporation of High-Level Features, Late Fusion and Discriminative Classification Methods. The development of robust multilingual speaker recognition systems will benefit the community through the elimination of fraud incurred by financial institutions and customers by enabling several person authentication applications such as: voice based signatures and document issuance; credit card verification by voice and secure over-the-phone financial transactions. The technology will also assist in the protection of the community and safeguard Australia by enabling the implementation of the following: suspect identification using voice print; national security measures for combating terrorism by using voice to locate and track terrorists; preemptive criminal activity counter-measures; surveillance and secure building access by voice.Read moreRead less
Robust speaker recognition with reduced utterance duration and intersession variability. The development of robust and accurate speaker recognition systems will enable secure person authentication in over-the-phone financial transactions and benefit the community through the elimination of identity fraud incurred by customers and financial institutions. The technology will also assist in safeguarding Australia by enabling the implementation of suspect identification using voice and security meas ....Robust speaker recognition with reduced utterance duration and intersession variability. The development of robust and accurate speaker recognition systems will enable secure person authentication in over-the-phone financial transactions and benefit the community through the elimination of identity fraud incurred by customers and financial institutions. The technology will also assist in safeguarding Australia by enabling the implementation of suspect identification using voice and security measures for combating terrorism by using voice to locate and track terrorists. Our research at QUT Speech Research Lab is at the forefront of development in this field and will provide Australia with a technological advantage in the rapidly evolving global market for speaker recognition technology for person authentication applications.Read moreRead less
Robust Automatic Speaker Diarisation of Audio Documents by Exploiting Prior Sources of Information. Speaker Diarisation, the task of determining who spoke when, is a technology fundamental in deriving intelligent information from audio and multimedia resources. The requirement for efficient and accurate Speaker Diarisation systems, portable across different domains is heightened by the explosive growth of audio and multimedia archives online and throughout the world. This research will provide t ....Robust Automatic Speaker Diarisation of Audio Documents by Exploiting Prior Sources of Information. Speaker Diarisation, the task of determining who spoke when, is a technology fundamental in deriving intelligent information from audio and multimedia resources. The requirement for efficient and accurate Speaker Diarisation systems, portable across different domains is heightened by the explosive growth of audio and multimedia archives online and throughout the world. This research will provide the foundation for a commercial service of automatic Speaker Diarisation to be developed, growing Australia's impact on the information and communications technology (ICT) sector. The outcome of this research will also assist in the tracking of terrorist and unlawful activity by enabling effective intelligence gathering from different audio sources.Read moreRead less
Unravelling A New Fatty Acid Pathway Involved In Neuroexocytosis And Memory
Funder
National Health and Medical Research Council
Funding Amount
$539,631.00
Summary
This proposal build on the establishment by our laboratory of the assay capable of detecting free fatty acids, with great accuracy and sensitivity. Using this assay we have uncovered a completely new pathway highlighting the production of saturated free fatty acids linked to learning and memory. We will fully define how this pathway is regulated in the brain.
Audio Visual Speech Recognition. Even though significant advances have been made in automatic speech recognition using acoustic information, the recognition accuracies are still poor in noisy and hostile environments such as in crowds, traffic, factory floors etc. In many of these applications visual information is or can easily be made available in addition to the audio. The aim of this project is to achieve an order of magnitude improvement in speech recognition accuracies in adverse environme ....Audio Visual Speech Recognition. Even though significant advances have been made in automatic speech recognition using acoustic information, the recognition accuracies are still poor in noisy and hostile environments such as in crowds, traffic, factory floors etc. In many of these applications visual information is or can easily be made available in addition to the audio. The aim of this project is to achieve an order of magnitude improvement in speech recognition accuracies in adverse environments by joint processing and modelling of the acoustic modality with visual information in the form of lip shapes and movements. The outcomes will be useful in human computer interaction in adverse environments as well as in the transcription and mining of multimedia data.
Read moreRead less
Cognitive Enhancement In Schizophrenia Via Selective Oestrogen Receptor Modulator.
Funder
National Health and Medical Research Council
Funding Amount
$396,380.00
Summary
Cognitive dysfunction in schizophrenia is resistant to treatment and related to poor community functioning and quality of life. In spite of the widely appreciated magnitude of the problem, there is still a critical gap in our knowledge concerning treatments to reverse these cognitive deficits. The proposed research is significant because it will clarify the role of hormones and genes in relation to cognitive deficits in schizophrenia and it may help patients improve their level of functioning.
Adaptive learning of spatiotemporal patterns: Development of multi-layer spiking neuron networks using Hebbian and competitive learning. The aim of this project is to develop a method for recognising patterns that change in time. The development of a reliable method that is fast and robust to noise will have wide application in many areas, especially computer speech recognition where timing plays a crucial role. Building-blocks similar to those in the brain (spiking neurons) will be used. Aut ....Adaptive learning of spatiotemporal patterns: Development of multi-layer spiking neuron networks using Hebbian and competitive learning. The aim of this project is to develop a method for recognising patterns that change in time. The development of a reliable method that is fast and robust to noise will have wide application in many areas, especially computer speech recognition where timing plays a crucial role. Building-blocks similar to those in the brain (spiking neurons) will be used. Automatic techniques will be used to teach groups of spiking neurons the differences between sequences of events by adjusting connections between them. The significance of this approach is that it captures information about timing that is missed in existing techniques.Read moreRead less