Speech Dataset

Chittagonian Dataset

High-Quality Chittagonian General Conversation and TTS Dataset for AI & Speech Models

Chittagonian Dataset

Overview

  • LanguageChittagonian
  • Dataset TypesGeneral Conversation, TTS
  • CountryBangladesh

Description

Unscripted, synthetic telephonic conversation between "speaker 1" and "speaker 2", Approx Audio Duration (Range) 5-15 Minutes.

Dataset Highlights

  • LanguageChittagonian
  • Sample Rate16 kHz / 44 kHz
  • Total Hours900 hrs
  • ChannelsMono
  • Duration Range5 - 15 mins

Use Cases

  • Politics
  • current affairs
  • local news
  • religion
  • economics and finance
  • and tourism

Sample Dataset Specs

Dataset TypeGeneral ConversationTTS
Sampling Rate44 kHz16 kHz
ChannelMonoMono
Total Hours100 hrs800 hrs

Key Benefits

  • Diverse & RepresentativeBroad speaker coverage for inclusive, real-world AI.
  • High-Quality AudioClean, balanced recordings for reliable model performance.
  • Ready for AI/MLStructured and annotated for seamless integration.
  • Scalable & FlexibleMultiple durations, speakers, and domains to fit your needs.

Technical Details

  • Recording PlatformMobile App
  • Audio Format.wav
  • Transcription Format.json
  • WER (%)5
  • Age Range18-50

Featured Clients

Trusted by leading AI teams and enterprises worldwide.

  • Microsoft
  • Amazon Web Services
  • Google

Can’t find what you are looking for?

New off-the-shelf datasets are being collected across all data types

Contact us now to let go of your audio/speech training data collection worries

  • This field is for validation purposes and should be left unchanged.
  • By registering, I agree with Shaip Privacy Policy and Terms of Service and provide my consent to receive B2B marketing communication from Shaip.