ARC Research Network for Enabling Human Communication. The Human Communication Network promotes interdisciplinary research in speech, language, and sound by and between humans and machines. The network connects leading and emerging researchers across disciplines, exploits previously unrecognised intersections, supports interdisciplinary graduate training and exchanges, provides database storage infrastructure, and consults with industry and government to set, not follow, research agendas. By ge ....ARC Research Network for Enabling Human Communication. The Human Communication Network promotes interdisciplinary research in speech, language, and sound by and between humans and machines. The network connects leading and emerging researchers across disciplines, exploits previously unrecognised intersections, supports interdisciplinary graduate training and exchanges, provides database storage infrastructure, and consults with industry and government to set, not follow, research agendas. By generating an explosion of new approaches and knowledge, the network will build Australia's reputation as a leader in communication science and technology via advances in automatic speech recognition, distress call monitoring, hearing prostheses, web interfaces, and data retrieval and data mining systems.Read moreRead less
Cognitive Modelling of Computer Games Pidgins. This project develops a pidgin language for use in computer games and models the way human users exploit such languages. It has implications also for computer assisted collaborative work and other educational or entertainment interactive environments like computer games. By developing mini-languages for games environments we can dramatically simplify the speech recognition problem and make recognition robust across different speech cultures and ba ....Cognitive Modelling of Computer Games Pidgins. This project develops a pidgin language for use in computer games and models the way human users exploit such languages. It has implications also for computer assisted collaborative work and other educational or entertainment interactive environments like computer games. By developing mini-languages for games environments we can dramatically simplify the speech recognition problem and make recognition robust across different speech cultures and backgrounds. We use protocol analysis and markup techniques for modelling dialogues between human player and reactive agents in computer games.Read moreRead less
Robust feature extraction for automatic speech recognition. Speech is perhaps the most natural and efficient mode of communication for humans. Therefore, it has always been a dream for many people to communicate with machines via speech. Significant advances have been made in the last five decades in the area of automatic speech recognition. Though the currently available speech recognisers work reasonably well in noise-free office environments, their performance deteriorates drastically when th ....Robust feature extraction for automatic speech recognition. Speech is perhaps the most natural and efficient mode of communication for humans. Therefore, it has always been a dream for many people to communicate with machines via speech. Significant advances have been made in the last five decades in the area of automatic speech recognition. Though the currently available speech recognisers work reasonably well in noise-free office environments, their performance deteriorates drastically when they are deployed in real-life situations due to the presence of background noise and other distortions. The problem of robust speech recognition will be researched in this project. Read moreRead less
Robust speech recognition in realistic hostile environments. Australia leads the world in the adoption of speech recognition technology but sadly lags in the development of the fundamental advances in the area. This research will help propel Australia to the forefront of new innovations in speech recognition technology and contributions to fundamental science. Our project will provide an excellent training ground for graduate students and researchers, with the real possibility of significant com ....Robust speech recognition in realistic hostile environments. Australia leads the world in the adoption of speech recognition technology but sadly lags in the development of the fundamental advances in the area. This research will help propel Australia to the forefront of new innovations in speech recognition technology and contributions to fundamental science. Our project will provide an excellent training ground for graduate students and researchers, with the real possibility of significant commercial benefit to the nation. The deployment of our system in the community will greatly enhance the defence and police forces ability for surveillance and security, and will provide new assistive aids to improve the quality of life and safety for the elderly and disabled.Read moreRead less
Enhanced Multilingual Speaker Recognition through the Incorporation of High-Level Features, Late Fusion and Discriminative Classification Methods. The development of robust multilingual speaker recognition systems will benefit the community through the elimination of fraud incurred by financial institutions and customers by enabling several person authentication applications such as: voice based signatures and document issuance; credit card verification by voice and secure over-the-phone financi ....Enhanced Multilingual Speaker Recognition through the Incorporation of High-Level Features, Late Fusion and Discriminative Classification Methods. The development of robust multilingual speaker recognition systems will benefit the community through the elimination of fraud incurred by financial institutions and customers by enabling several person authentication applications such as: voice based signatures and document issuance; credit card verification by voice and secure over-the-phone financial transactions. The technology will also assist in the protection of the community and safeguard Australia by enabling the implementation of the following: suspect identification using voice print; national security measures for combating terrorism by using voice to locate and track terrorists; preemptive criminal activity counter-measures; surveillance and secure building access by voice.Read moreRead less
Determinants of Audio-Visual effects in degraded and non-degraded speech. Seeing a speaker's face can affect the perception of their speech in a number of ways. This project proposes a detailed comparison of factors that affect Audio-Visual (AV) facilitation of degraded speech detection and identification. Detection-based tasks should be more sensitive to signal based correlations whereas identification-based effects more sensitive to complementary information. The significance of the current pr ....Determinants of Audio-Visual effects in degraded and non-degraded speech. Seeing a speaker's face can affect the perception of their speech in a number of ways. This project proposes a detailed comparison of factors that affect Audio-Visual (AV) facilitation of degraded speech detection and identification. Detection-based tasks should be more sensitive to signal based correlations whereas identification-based effects more sensitive to complementary information. The significance of the current proposal is that it offers both a strategy and a connected series of experiments for determining key behavioural constraints on AV speech integration. Understanding AV interactions will build links between neurophysiological processes and coherent perception and have important implications for AV application.Read moreRead less
Robust speaker recognition with reduced utterance duration and intersession variability. The development of robust and accurate speaker recognition systems will enable secure person authentication in over-the-phone financial transactions and benefit the community through the elimination of identity fraud incurred by customers and financial institutions. The technology will also assist in safeguarding Australia by enabling the implementation of suspect identification using voice and security meas ....Robust speaker recognition with reduced utterance duration and intersession variability. The development of robust and accurate speaker recognition systems will enable secure person authentication in over-the-phone financial transactions and benefit the community through the elimination of identity fraud incurred by customers and financial institutions. The technology will also assist in safeguarding Australia by enabling the implementation of suspect identification using voice and security measures for combating terrorism by using voice to locate and track terrorists. Our research at QUT Speech Research Lab is at the forefront of development in this field and will provide Australia with a technological advantage in the rapidly evolving global market for speaker recognition technology for person authentication applications.Read moreRead less
Visual Solutions for Automated Translation Between Spoken and Signed Languages. We propose to build a robust visual speech recognition system that analyzes images of spoken language and achieves a recognition of the utterances with at least human expert recognition rates. This visual speech recognition system will then be integrated with our existing gesture recognition system to improve performance, just as humans combine visual and audio data for language understanding. The result will be a sy ....Visual Solutions for Automated Translation Between Spoken and Signed Languages. We propose to build a robust visual speech recognition system that analyzes images of spoken language and achieves a recognition of the utterances with at least human expert recognition rates. This visual speech recognition system will then be integrated with our existing gesture recognition system to improve performance, just as humans combine visual and audio data for language understanding. The result will be a system providing translation between English and the Australian sign language Auslan in a practical application domain. Significantly, our work will provide insights into the cognitive models of neural activity linking language and gesture.Read moreRead less
Robust Automatic Speaker Diarisation of Audio Documents by Exploiting Prior Sources of Information. Speaker Diarisation, the task of determining who spoke when, is a technology fundamental in deriving intelligent information from audio and multimedia resources. The requirement for efficient and accurate Speaker Diarisation systems, portable across different domains is heightened by the explosive growth of audio and multimedia archives online and throughout the world. This research will provide t ....Robust Automatic Speaker Diarisation of Audio Documents by Exploiting Prior Sources of Information. Speaker Diarisation, the task of determining who spoke when, is a technology fundamental in deriving intelligent information from audio and multimedia resources. The requirement for efficient and accurate Speaker Diarisation systems, portable across different domains is heightened by the explosive growth of audio and multimedia archives online and throughout the world. This research will provide the foundation for a commercial service of automatic Speaker Diarisation to be developed, growing Australia's impact on the information and communications technology (ICT) sector. The outcome of this research will also assist in the tracking of terrorist and unlawful activity by enabling effective intelligence gathering from different audio sources.Read moreRead less
ARC Communications Research Network. Building on a strong platform of existing research excellence, the Aim of the Network is to facilitate nation-wide collaborative research, promoting four intersecting research Themes: Mobile and Wireless Communications, Rural Communications, Broadband and Optical Networks, and Fundamentals of Emerging Media. Each Theme is formulated to drive multidisciplinary, innovative research as well as inspire new collaborative initiatives. Four Programs encapsulate the ....ARC Communications Research Network. Building on a strong platform of existing research excellence, the Aim of the Network is to facilitate nation-wide collaborative research, promoting four intersecting research Themes: Mobile and Wireless Communications, Rural Communications, Broadband and Optical Networks, and Fundamentals of Emerging Media. Each Theme is formulated to drive multidisciplinary, innovative research as well as inspire new collaborative initiatives. Four Programs encapsulate the core activities of the Network: Researcher Mobility, Workshops and Conferences, Postgraduate Education, and Knowledge Management Systems. The Network is expected to add significant value to pre-existing investments and raise the profile of Australian telecommunications research.Read moreRead less