CertKeen

Google Cloud Professional Cloud Architect · Free practice question 2 of 12

Chirp speech model for call transcription

Dunwell Utilities records 40,000 customer service calls a month and wants accurate text transcripts in several languages for quality review and search. The deciding constraint is a managed Google model rather than training its own. Which service should the architect recommend?

  1. A.Text-to-Speech with a Chirp HD voice
  2. B.Speech-to-Text with a Chirp model, using batch recognition
  3. C.Cloud Translation API applied directly to the audio files
  4. D.A custom acoustic model trained from scratch on Compute Engine GPUs with the call recordings
Show answer and explanation

Correct answer: B. Speech-to-Text with a Chirp model, using batch recognition

Why: Speech-to-Text converts audio to text, and its Chirp models are Google's large multilingual speech models; batch recognition suits recorded calls. Translation works on text, not audio. Text-to-Speech generates audio from text rather than transcribing it, and training an acoustic model from scratch contradicts the constraint.

More free Google Cloud Professional Cloud Architect questions