Creating Subway-Like Speech by Mixing Clean Voice and Transit Noise
If a speech application needs audio that sounds like it was captured on a crowded train platform, one practical synthesis method is to combine clean voice recordings with background transit noise. The mixed audio can serve as a stand-in for speech recorded in that noisy setting.
0
1
Tags
Machine Learning
Deep Learning
Supervised Learning
Dive into Deep Learning @ D2L
Data Science
Machine Learning Strategy
Machine Learning Yearning @ DeepLearning.AI
Related
Creating Subway-Like Speech by Mixing Clean Voice and Transit Noise
Adding Artificial Blur to Training Photos
A Gap Between Human and Machine Judgments of Synthetic Data
When Synthetic Data Becomes Useful
What does artificial data synthesis help you build when your development set is missing important cases?
True or False: If your validation data contains an important rare pattern, generating synthetic examples can help enlarge the training set so it better covers that pattern.
Artificial data synthesis can help create a _____ that better matches the validation set.
What is the main advantage of synthetic data when your training set does not match the dev set?
Artificially generated examples always match the dev set’s real-world distribution exactly.
Artificially generated examples can help create a _____ dataset that still resembles the development set.
Match each synthetic-data situation to the real-world factor it is meant to imitate.
Order the reasoning steps for deciding whether generated data can help match a validation distribution.
When is synthetic data most useful for matching a development set?
Synthetic examples can help narrow the difference between training data and development data distributions.
There are several _____ in which artificial data generation can produce a large dataset that closely matches the development set.
Match each concept in synthetic-data design to its best description.
Order the steps for creating synthetic office-call audio to resemble a noisy support-center dev set.
When is synthetic data useful for matching a development set?
When Synthetic Data Is Worth Building for a Narrow Validation Set
What should synthesized training data achieve when it is built to mirror a dev set?
Learn After
What data combination can make speech sound as if it was recorded in a loud vehicle?
Synthetic driving audio can reduce the amount of real recording data needed.
Mixing clean speech with _____ can create a subway-platform recording.
Match each audio item in a synthesized announcement task to its role.
Order the steps for creating speech data that sounds like it was recorded in a subway station.
Create synthetic transit speech from two audio sources
Create training audio that sounds like speech heard inside a vehicle
What audio sources are combined to create speech that sounds noisy?
If you mix a clean voice recording with city traffic noise, what should the result sound like?
Can background engine noise by itself provide examples of spoken sentences for a speech recognizer?