QuestionQ58

Automating and orchestrating ML pipelines

Your company retains many audio recordings of calls made to its customer call center in an on-premises database. Each recording is in WAV format and lasts approximately 5 minutes. You need to analyze the audio files for customer sentiment. You plan to use the Speech-to-Text API and want the most efficient approach. What should you do?

Explanation

Five-minute recordings require asynchronous speech:longrunningrecognize, because synchronous recognition is limited to audio of one minute or less. Storing the audio in Cloud Storage lets Speech-to-Text access it without repeatedly sending local file content. The Natural Language API's analyzeSentiment method is the prebuilt service for determining sentiment from the resulting transcription, so creating and invoking an AutoML sentiment model is unnecessary. A Cloud Function can automate that downstream analysis.

Learn more

Community Discussion

No comments yet. Be the first to start the discussion!