Practice

(1/5)

1. What is the main purpose of text-to-speech (TTS) technology?

easy

A. To summarize long documents automatically

B. To translate text from one language to another

C. To detect emotions in spoken language

D. To convert written text into spoken audio

Solution

Step 1: Understand the function of TTS
Text-to-speech technology changes written words into sound that can be heard.
Step 2: Compare options with TTS purpose
Only To convert written text into spoken audio describes converting text to speech, which matches TTS.
Final Answer:
To convert written text into spoken audio -> Option D
Quick Check:
TTS = convert text to speech [OK]

Hint: Remember TTS means text becomes speech [OK]

Common Mistakes:

Confusing TTS with translation
Thinking TTS summarizes text
Mixing TTS with emotion detection

2. Which Python library is commonly used for simple text-to-speech conversion?

easy

A. Pandas

B. gTTS

C. Matplotlib

D. NumPy

Solution

Step 1: Identify libraries related to TTS
gTTS is a Python library designed for text-to-speech conversion.
Step 2: Eliminate unrelated libraries
NumPy, Matplotlib, and Pandas are for math, plotting, and data, not TTS.
Final Answer:
gTTS -> Option B
Quick Check:
gTTS = text-to-speech library [OK]

Hint: gTTS stands for Google Text-to-Speech [OK]

Common Mistakes:

Choosing data or plotting libraries by mistake
Confusing gTTS with general Python packages
Assuming TTS needs complex libraries always

3. What will the following Python code output?

from gtts import gTTS
text = 'Hello world'
tts = gTTS(text)
tts.save('hello.mp3')
print('Audio saved')

medium

A. An audio file named 'hello.mp3' is created and 'Audio saved' is printed

B. The text 'Hello world' is printed on screen

C. A syntax error occurs due to missing language parameter

D. Nothing happens because gTTS requires internet connection

Solution

Step 1: Analyze the code steps
The code imports gTTS, creates speech from 'Hello world', saves it as 'hello.mp3', then prints a message.
Step 2: Check for errors or missing parts
gTTS defaults to English if no language is given, so no syntax error occurs. Internet is needed but code runs assuming connection.
Final Answer:
An audio file named 'hello.mp3' is created and 'Audio saved' is printed -> Option A
Quick Check:
Code saves audio and prints message [OK]

Hint: gTTS saves audio file and prints confirmation [OK]

Common Mistakes:

Thinking language parameter is mandatory
Assuming print outputs the text spoken
Ignoring that gTTS needs internet but code runs

4. Identify the error in this text-to-speech code snippet:

from gtts import gTTS
tts = gTTS('Hello')
tts.save()

medium

A. Missing filename argument in save() method

B. gTTS requires language parameter in constructor

C. Text argument should be a list, not a string

D. gTTS cannot be imported directly

Solution

Step 1: Check gTTS usage
gTTS constructor accepts text string; language is optional. So no error there.
Step 2: Check save() method
save() requires a filename string argument to save the audio file. Missing argument causes error.
Final Answer:
Missing filename argument in save() method -> Option A
Quick Check:
save() needs filename [OK]

Hint: save() always needs a filename string [OK]

Common Mistakes:

Assuming language is always required
Thinking text must be a list
Believing import statement is wrong

5. You want to create a text-to-speech system that can speak multiple languages based on user input. Which approach is best?

hard

A. Use gTTS without language parameter and rely on default English

B. Manually translate text first, then use gTTS with fixed language

C. Use gTTS with a dynamic language parameter set from user input

D. Use a single pre-recorded audio file for all languages

Solution

Step 1: Understand multilingual TTS needs
The system must speak different languages based on user choice, so language must be flexible.
Step 2: Evaluate options for language flexibility
Use gTTS with a dynamic language parameter set from user input sets language dynamically in gTTS, allowing correct speech for each language. Others fix language or use static audio, which won't adapt.
Final Answer:
Use gTTS with a dynamic language parameter set from user input -> Option C
Quick Check:
Dynamic language parameter enables multilingual TTS [OK]

Hint: Set language parameter dynamically for multilingual speech [OK]

Common Mistakes:

Ignoring language parameter flexibility
Assuming default English works for all
Using static audio files for dynamic text

Epoch	Loss ↓	Accuracy ↑	Observation
1	2.5	0.30	Model starts learning basic phoneme to sound mapping
5	1.2	0.55	Improved clarity in generated mel-spectrograms
10	0.7	0.75	Neural vocoder produces more natural waveforms
15	0.4	0.85	Speech sounds clear and intelligible
20	0.25	0.92	Model converges with high quality speech output

Text-to-speech generation in Prompt Engineering / GenAI - Model Pipeline Trace

Start learning this pattern below

Practice

Solution

Step 1: Understand the function of TTS

Step 2: Compare options with TTS purpose

Final Answer:

Quick Check:

Solution

Step 1: Identify libraries related to TTS

Step 2: Eliminate unrelated libraries

Final Answer:

Quick Check:

Solution

Step 1: Analyze the code steps

Step 2: Check for errors or missing parts

Final Answer:

Quick Check:

Solution

Step 1: Check gTTS usage

Step 2: Check save() method

Final Answer:

Quick Check:

Solution

Step 1: Understand multilingual TTS needs

Step 2: Evaluate options for language flexibility

Final Answer:

Quick Check: