You should use Real-Time Speech-to-Text (Streaming STT) for this project. Why Real-Time STT is Required Low Latency: It transcribes audio within seconds. Continuous Input: It processes live, ongoing audio streams. Live Monitoring: Supervisors need immediate text to flag issues. Reference: https://learn.microsoft.com/en-us/azure/ai-services/speech-service/get-started-speech-to-text