Speech Dataset

Vietnamese Dataset

3 licensable sets · 2 dataset types · 1 country

3 licensable Vietnamese speech sets: 7,700 hours across 2 dataset types, sampled at 16 kHz in mixed mono/stereo, dual and mono channel.

Vietnamese Dataset

Overview

  • LanguageVietnamese
  • Dataset TypesGeneral Conversation, Scripted Monologue
  • CountryVietnam

Dataset Highlights

  • LanguageVietnamese
  • Sample Rate16 kHz
  • Total Hours7,700 hrs
  • ChannelsMono & Stereo (mixed) / Dual / Mono
Available sets

What you can license today

Volumes are as reported by the sourcing partner; confirmed at scoping. Pricing is quoted per audio hour.

Language & Accent · CountryDataset TypeVolumeAudio ChannelAudio Frequency
VietnameseVietnamGeneral Conversation5,500 hoursMono & Stereo (mixed)—
VietnameseVietnamGeneral Conversation1,000 hoursDual16 kHz+
VietnameseVietnamScripted Monologue1,200 hoursMono16 kHz+

Use Cases

  • ASR (Automatic Speech Recognition)
  • Conversational AI
  • Language Modelling

Key Benefits

  • Diverse & RepresentativeBroad speaker coverage for inclusive, real-world AI.
  • High-Quality AudioClean, balanced recordings for reliable model performance.
  • Ready for AI/MLStructured and annotated for seamless integration.
  • Scalable & FlexibleMultiple durations, speakers, and domains to fit your needs.

Featured Clients

Trusted by leading AI teams and enterprises worldwide.

  • Microsoft
  • Amazon Web Services
  • Google

Can't find the Vietnamese set you need?

Tell us the dataset type, volume and recording conditions — we scope custom collection in this language too.

  • This field is for validation purposes and should be left unchanged.
  • By registering, I agree with Shaip Privacy Policy and Terms of Service and provide my consent to receive B2B marketing communication from Shaip.