Mel-frequency cepstrum
In sound processing, the mel-frequency cepstrum (MFC) is a representation of the short-term power spectrum of a sound, based on a linear cosine transform of a log power spectrum on a nonlinear mel scale of frequency.
Mel-frequency cepstrum - Wikipedia Jump to content From Wikipedia, the free encyclopedia Signal representation used in automatic speech recognition In sound processing , the mel-frequency cepstrum ( MFC ) is a representation of the short-term power spectrum of a sound, based on a linear cosine transform of a log power spectrum on a nonlinear mel scale of frequency. Mel-frequency cepstral coefficients ( MFCCs ) are coefficients that collectively make up an MFC. [ 1 ] They are derived from a type of cepstral representation of the audio clip (a nonlinear "spectrum-of-a-spectrum"). The difference
Explore this link on the map →related reading
- Mel scale - Wikipediaen.wikipedia.org
- How does Audio Fingerprinting work - Emysoundemysound.com
- Audio Deep Learning Made Simple: Automatic Speech Recognition (ASR), How it Works | Towards Data Sciencetowardsdatascience.com
- Diffusion is spectral autoregression – Sander Dielemansander.ai
- TF Representations and Masking — Open-Source Tools & Data for Music Source Separationsource-separation.github.io
- Sequence Modeling with CTCdistill.pub
- Audio time stretching and pitch scaling - Wikipediaen.wikipedia.org
- Generating music in the waveform domain – Sander Dielemansander.ai
- Short-time Fourier transform - Wikipediaen.wikipedia.org
- Prof. Dr. M. R. Schroederweb.archive.org
- Formant - Wikipediaen.wikipedia.org
- Discrete cosine transform - Wikipediam.wikipedia.org