Speech Dataset

Cantonese Dataset

2 licensable sets · 1 dataset type · 2 countries

2 licensable Cantonese speech sets: 110 hours across 1 dataset type, sampled at 16 kHz in mono channel.

Cantonese Dataset

Overview

  • LanguageCantonese
  • Dataset TypesGeneral Conversation
  • CountryChina, Hong Kong

Dataset Highlights

  • LanguageCantonese
  • Sample Rate16 kHz
  • Total Hours110 hrs
  • ChannelsMono
Available sets

What you can license today

Volumes are as reported by the sourcing partner; confirmed at scoping. Pricing is quoted per audio hour.

Language & Accent · CountryDataset TypeVolumeAudio ChannelAudio Frequency
CantoneseHong KongGeneral Conversation60 hoursMono16 kHz
CantoneseChinaGeneral Conversation50 hoursMono16 kHz

Use Cases

  • ASR (Automatic Speech Recognition)
  • Conversational AI
  • Language Modelling

Key Benefits

  • Diverse & RepresentativeBroad speaker coverage for inclusive, real-world AI.
  • High-Quality AudioClean, balanced recordings for reliable model performance.
  • Ready for AI/MLStructured and annotated for seamless integration.
  • Scalable & FlexibleMultiple durations, speakers, and domains to fit your needs.

Featured Clients

Trusted by leading AI teams and enterprises worldwide.

  • Microsoft
  • Amazon Web Services
  • Google

Can't find the Cantonese set you need?

Tell us the dataset type, volume and recording conditions — we scope custom collection in this language too.

  • This field is for validation purposes and should be left unchanged.
  • By registering, I agree with Shaip Privacy Policy and Terms of Service and provide my consent to receive B2B marketing communication from Shaip.