- Remaps frequency so small changes at low frequencies matter more than equal-sized changes at high frequencies.
- Practical default for recognition and synthesis: compact, while keeping the parts of the signal that matter for speech.
- Whisper mental model: PCM → window → FFT → mel filters + log compression → transformer.
Connections
No outgoing connections.