Noisy Speech Datasets for Model Robustness

Noisy speech datasets recorded in cafes, streets, offices, and vehicles. Real-world audio for training ASR models that stay robust in production noise.

  • Real-World Audio: Cafes, Streets, Offices, Vehicles
  • Datasets with Signal-to-Noise Ratio (SNR) Metadata
  • Paired Clean/Noisy Datasets Available

Dataset information

Creator
YPAI
Formats
audio/wav, audio/flac

Check availability and fit

Review the current catalog, then request the license, provenance, metadata, and delivery details for the dataset you are evaluating.

Access Noise Library