speech.zone

Simon October 11, 2014

Windowing

slownormalfast

When we say that a signal is non-stationary we mean that its properties, such as the spectrum, change over time. To analyse signals like this, we need to first assume that these properties do not change over some short period of time, called the frame. We can then analyse individual frames of the signal, one at a time – we perform a short-term analysis. When extracting a frame, its important to apply a window with tapered edges, to remove discontinuities at the start and end of the frame. Here we see why that is, and what would happen if we forgot to apply a tapered window.

Try it for yourself – here are the materials to download:

A frame of speech extracted from a larger waveform: framed
The same frame of speech, after a tapered window has been applied: framed_windowed

Filed Under: Signals Tagged With: Short-term analysis, video, Wavesurfer

Simon November 1, 2022

Bitrate

The bitrate (or bit rate) of a signal is the number of bits required to store, or transmit, 1 s of that signal. A bit is a binary number: either 0 or 1. Let’s calculate the bitrate of a digital waveform. First you should revise the concepts of sampling and quantisation from this module of the […]

Filed Under: Signals Tagged With: Digital signal

Simon October 25, 2014

A super-simple speech recogniser

We make what is possibly the world’s simplest speech recognition system. It can only recognise two different words, but will help you understand the basic idea of pattern recognition using template matching. The templates are just pre-recorded words, with known labels. The features extracted are just two formant frequencies in the middle of the word, […]

Filed Under: Recognition Tagged With: Classification, equations, Wavesurfer

Simon February 1, 2015

Autocorrelation for estimating F0

Most methods for estimating F0 start from autocorrelation. The idea is pretty simple: we are just looking for a repeating pattern in the waveform, which corresponds to the periodic vocal fold activity. For some waveforms, it might be possible to do that directly in the time domain, but in general that doesn’t work very well. […]

Filed Under: Signals Tagged With: spreadsheet, video

Simon October 11, 2014

Pipeline architecture for TTS

Most text-to-speech systems split the problem into two main stages. The first stage is called the front end and contains many separate processes which gradually build up a linguistic specification from the input text. The second stage typically uses language-independent techniques (although they still require a language-specific speech corpus) to generate a waveform. Here we see those two […]

Filed Under: Synthesis Tagged With: front end, video, waveform generation

Simon October 31, 2015

The speed of sound

At the Parque de las Ciencias in Granada, Spain there is this long tube, open at the end nearest you and closed at the far end. We can calculate the length of this tube just from the audio recording, because we know the speed of sound. Here’s the waveform of part of the recording, showing […]

Filed Under: Signals Tagged With: video, Wavesurfer

Simon October 11, 2014

Classification and regression trees (CART)

A quick introduction to a very simple but widely-applicable model that can perform classification (predicting a discrete label) or regression (predicting a continuous value). The tree is learned from labelled data, using supervised learning. Before watching this video, you might want to check that you understand what Entropy is.

Filed Under: Models Tagged With: Classification, Decision tree, Learning decision trees, supervised learning, video

Simon October 11, 2014

Aliasing

In sampling and quantisation we saw that sampling a signal at a fixed rate means that there is an upper limit on the frequencies that can be represented. This limit is called the Nyquist frequency. Before sampling a signal, we must remove all energy above the Nyquist frequency, and here we will see what would […]

Filed Under: Signals Tagged With: Digital signal, video

Simon October 11, 2014

Sampling and quantisation

Is digital better than analogue? Here we discover that there are limitations when storing waveforms digitally. We learn that the consequence of sampling at a fixed rate is an upper limit on the frequencies that can be represented, called the Nyquist frequency. In addition to the limitations of sampling, storing each sample of the waveform as a […]

Filed Under: Signals Tagged With: Digital signal, video, Wavesurfer

Simon November 15, 2014

Token passing

Token passing is a really nice way to understand (and even to implement) Viterbi search for Hidden Markov Models. Here we see token passing in action, and you can look at the spreadsheet to see the calculations. To keep things simple, we are ignoring transition probabilities in this example. It would be simple to add them […]

Filed Under: Models, Recognition Tagged With: HMMs, spreadsheet, video

Simon October 12, 2014

Entropy: understanding the equation

The equation for entropy is very often presented in textbooks without much explanation, other than to say it has the desired properties. Here, I attempt an informal derivation of the equation starting from uniform probability distributions. A good way to think about information is in terms of sending messages. In the video, we send messages […]

Filed Under: Probability Tagged With: entropy, equations, video

Simon February 6, 2012

My inaugural lecture

I talk about how speech synthesis works, in what I hope is a non-technical and accessible way, and finish off with an application of speech synthesis that gives personalised voices to people who are losing the ability to speak. I also try to mention bicycles as many times as possible. For a more up-to-date, slightly more technical, […]

Filed Under: Synthesis Tagged With: lecture, video

Windowing

Bitrate

A super-simple speech recogniser

Autocorrelation for estimating F0

Pipeline architecture for TTS

The speed of sound

Classification and regression trees (CART)

Aliasing

Sampling and quantisation

Token passing

Entropy: understanding the equation

My inaugural lecture

Search this site

Posts

Latest Activity