Speech Dataset

Spanish Dataset

High-Quality Spanish Dataset for AI & Speech Models

Spanish Dataset

Overview

  • LanguageSpanish
  • Dataset TypesCall-Center, Podcast, Scripted Monologue
  • CountrySpain

Description

This dataset includes unscripted synthetic agent–customer telephonic conversations (5–15 minutes), and singing audio with transcriptions, providing diverse speech data for training and evaluating speech and language technologies.

Dataset Highlights

  • LanguageSpanish
  • Sample Rate24 kHz / 48 kHz / 8 kHz
  • Total Hours3287:25:30
  • Speakers2,242+
  • ChannelsDual / Mono
  • Duration Range5 - 15 mins

Use Cases

  • ASR (Automatic Speech Recognition)
  • Virtual Assistants & Chatbots
  • Conversational AI
  • Speech Analytics
  • TTS (Text-to-Speech)
  • Language Modelling

Sample Dataset Specs

Dataset TypeCall CenterCall CenterMusicScripted MonologueScripted Monologue
Sampling Rate8 kHz8 kHz48 kHz48 kHz24 kHz
Speakers2 Speakers2 SpeakersSingle SpeakerSingle SpeakerSingle Speaker
ChannelDualMonoMonoMonoMono
Total Hours61:08:06600:00:0005:17:241,521:00:001,100:00:00
Total # of SpeakersOn RequestOn Request382,204On Request

Key Benefits

  • Diverse & RepresentativeBroad speaker coverage for inclusive, real-world AI.
  • High-Quality AudioClean, balanced recordings for reliable model performance.
  • Ready for AI/MLStructured and annotated for seamless integration.
  • Scalable & FlexibleMultiple durations, speakers, and domains to fit your needs.

Featured Clients

Trusted by leading AI teams and enterprises worldwide.

  • Microsoft
  • Amazon Web Services
  • Google

Can’t find what you are looking for?

New off-the-shelf datasets are being collected across all data types

Contact us now to let go of your audio/speech training data collection worries

  • This field is for validation purposes and should be left unchanged.
  • By registering, I agree with Shaip Privacy Policy and Terms of Service and provide my consent to receive B2B marketing communication from Shaip.