Audio-Visual Speech Recognition
Combines acoustic speech with lip and facial motion to improve speech-recognition robustness.
Technical Context
Combines acoustic speech with lip and facial motion to improve speech-recognition robustness.
This concept belongs to the speech and audio pipeline from acoustic representation to decision-making.