Speech Recognition
Speech recognition: audio waveform → text. Steps: features (MFCCs or learned), acoustic model, language model (today often end-to-end nets). Noise, accents, and overlapping talk still hurt. Not the same as speaker identification (who spoke).
Output — idea / do / check for Speech Recognition. Fill those three in the viva.
Exam tip
ASR vs speaker ID.