Svenska Dataset
High-Quality Swedish Call-Center, and Media (Podcast) Dataset for AI & Speech Models
This dataset includes unscripted synthetic agent–customer telephonic conversations (5–15 minutes) and licensable public domain audio or video files, such as interviews and podcasts with 1 to 5 participants (15–60 minutes).
| Dataset Type | Call Center | Call Center | Media Data |
|---|---|---|---|
| Sampling Rate | 8 kHz | 8 kHz | 16 kHz |
| Speakers | 2 Speakers | 2 Speakers | Multiple Speakers |
| Channel | Dual | Mono | Mono |
| Total Hours | 309:00:45 | 500:00:00 | 278:10:43 |
| Total # of Speakers | 2,568 | - | 718 |
Trusted by leading AI teams and enterprises worldwide.
New off-the-shelf datasets are being collected across all data types
Contact us now to let go of your audio/speech training data collection worries
We use cookies to improve your experience on our site. By using our site, you consent to cookies.
Manage your cookie preferences below:
Essential cookies enable basic functions and are necessary for the proper function of the website.
Google Tag Manager simplifies the management of marketing tags on your website without code changes.
Statistics cookies collect information anonymously. This information helps us understand how visitors use our website.
Google Analytics is a powerful tool that tracks and analyzes website traffic for informed marketing decisions.
Service URL: policies.google.com (opens in a new window)
Marketing cookies are used to follow visitors to websites. The intention is to show ads that are relevant and engaging to the individual user.
Google Ads is an online advertising platform that enables businesses to create targeted ads displayed on Google search results and partner sites.
Service URL: policies.google.com (opens in a new window)
You can find more information in our Cookie Policy and Privacy Policy.