Series
AY 2005
Applied Acoustics
This course introduces fundamental knowledge about speech analysis, speech coding, speech recognition, speech synthesis, speech signal processing, spoken dialogue systems, etc. It gathers a wide variety of key ideas on the techniques involved in these topics, a domain where Japan has been a leading country. Many applications such as speech and music information compression techniques in mobile phones, MD or MP3 are already well established, and speech recognition techniques or speech synthesis systems, although they have not yet reached human’s ability, already are very advanced information processing technologies. These topics involve algorithms and basic ideas of spectral analysis, pattern recognition, stochastic modeling, statistical training, optimization, etc. This course aims at acquiring the fundamental concepts and knowledge of these techniques.
Content List
Resources
Available
Available
#1 A: Phonetics [What is speech?]
Lecturer | Shigeki Sagayama
Resources
Available
Available
Resources
Available
Available
#2 B: Speech spectrum analysis [Basics of speech signal analysis]
Lecturer | Shigeki Sagayama
Resources
Available
Available
Resources
Available
Available
#3 C: LPC analysis [All pole speech modeling]
Lecturer | Shigeki Sagayama
Resources
Available
Available
Resources
Available
Available
#4 D: Clustering and Vector Quantization [Basis of speech coding]
Lecturer | Shigeki Sagayama
Resources
Available
Available
Resources
Available
Available
#5 E: Non-linear time warping [Basis of speech recognition]
Lecturer | Shigeki Sagayama
Resources
Available
Available
Resources
Available
Available
#6 F: Hidden Markov Models(HMM) [Phoneme modeling for speech recognition]
Lecturer | Shigeki Sagayama
Resources
Available
Available
Resources
Available
Available
#7 G: Speech recognition
Lecturer | Shigeki Sagayama
Resources
Available
Available
Resources
Available
Available
#8 H: Speech Synthesis
Lecturer | Shigeki Sagayama
Resources
Available
Available