Skip to main content
zenodoopen

Figure 2. MFCC Block Diagram Step 1-Development an Automatic Speech to Facial Animation Conversion for Improve Deaf Lives

<p>Mel Frequency Cepstral Coefficients (MFCC) are coefficients that represent audio, based on<br> perception. It is derived from the Fourier Transform (FFT) or the Discrete Cosine Transform (DCT)<br> of the audio clip. The basic difference between the FFT/DCT and the MFCC is that in the MFCC,<br> the frequency bands are positioned logarithmically (on the Mel scale) which approximates the<br> human auditory system&#39;s response more closely than the linearly spaced frequency bands of FFT or<br> DCT. This allows for better processing of data. The main purpose of the MFCC processor is to<br> mimic the behavior of the human ears. Overall the MFCC process has 5 steps that show in figure 2.</p>

ShareScore

40/100

Overall dataset sharing score

Score breakdown

These five areas show where the dataset supports — or may limit — practical reuse.

Stewardship
8
Harmonization
4
Access
20
Reuse readiness
8
Engagement
0

Topics