Swedish Speech Dataset For Speech and Voice Recognition Models

We provide Swedish Speech Dataset for training and testing Swedish speech/voice recognition algorithms and ASR models. Our transcribed NLP Dataset is perfect for speech-to-text and ASR models for Swedish language.

We have multiple datasets that you can choose from: transcribed spontaneous speech data with one or two people speaking or scripted monologues.


German Autolabs

Swedish Voice Dataset

Swedish voice dataset, Swedish voice recognition dataset, Swedish voice recognition data, Swedish voice recognition, Swedish voice data transcribed for ASR, Swedish voice data transcription for ASR, Swedish voice data, Swedish voice training dataset, Swedish voice testing dataset, Swedish voice recognition solution, Swedish voice recognition AI, Swedish voice recognition algorithms, common voice dataset for Swedish language, Swedish voices dataset, Swedish voice dataset kaggle, dataset for Swedish voice recognition, Swedish voice command dataset, Swedish voice dataset machine learning, Swedish voice emotion recognition dataset, Swedish voice data set

Swedish NLP Datasets

Swedish nlp datasets, Swedish dataset for nlp, Swedish natural language processing data sets, Swedish nlp dataset, named entity recognition dataset, Swedish nlp projects kaggle, kaggle nlp datasets, Swedish natural language dataset, Swedish text dataset for nlp, Swedish nlp training dataset, Swedish nlp datasets kaggle, Swedish dataset nlp, Swedish nlp classification datasets, Swedish datasets for nlp projects, Swedish nlp conversation dataset, Swedish natural language inference dataset, coreference resolution dataset, Swedish natural language datasets, Swedish nlp sentiment analysis dataset, Swedish datasets for natural language processing, Swedish nlp small dataset, Swedish dataset for nlp sentiment analysis, Swedish data sets for nlp

Swedish Language Speech Recognition Models

speech recognition neural network, speech to text neural network, voice recognition neural network, convolutional neural network speech recognition, Swedish voice activity detection model, Swedish asr language model, attention based models for Swedish speech recognition, Swedish language model in speech recognition, rnn speech recognition, lstm speech recognition, Swedish voice recognition model, Swedish speech recognition language model, Swedish language model speech recognition, Swedish language speaker identification model, best Swedish speech recognition models, Swedish speaker recognition model, Swedish speech to text deep learning model, Swedish asr acoustic model, Swedish language model for speech recognition, best asr models for Swedish language, Swedish speech recognition pretrained model, Swedish asr model, neural network for voice recognition, neural network speech to text, Swedish acoustic model in speech recognition, speech recognition with deep recurrent neural networks, Swedish acoustic model speech recognition, recurrent neural network speech recognition, Swedish speech to text ai model

Swedish Language Speech to Text Dataset

Swedish language speech to text dataset, Swedish language dataset for speech to text, Swedish language asr dataset, Swedish language machine translation dataset, Swedish language text to speech dataset, Swedish language speech to text dataset kaggle, Swedish language speech to text kaggle, Swedish language dataset for speech to text, Swedish language text to speech dataset, Swedish language text to speech data, Swedish language data transcription

Datasets for your speech recognition solution in Swedish

Improve your Swedish automatic speech recognition models or deploy new models in days using our speech and voice recognition dataset. The Swedish datasets you can choose from are scripted and non scripted recordings with one or two people speaking. Tell us what data you need and we will include only the data that fits your use case and needs, whether that is specific background noise levels, speakers from certain regions, speakers of specific age groups, gender, or nativitiy.

We can provide you with thousands of hours of speech recorded by tens of thousands unique speakers. With our high-quality training datasets, you can gain competitive advantage over your competitors, reduce time to market, and improve word error rate of your models.


Our speech recognition datasets in Swedish consists of native and non-native speakers from the following regions:
Swedish language: native SE, FI, and non-native.

Speech recognition data specifications

The Swedish The datasets contain transcribed and segmented audio clips of people talking about various topics or reading sentences, with up to two hours of speech per person. The speech is captured using mobile phones and laptops from a diverse crowd of speakers representing all ages and backgrounds. Because of that, the dataset is perfect for ASR and voice assistant use cases using mobile devices.

Recordings vary in length depending on type of recording. Scripted speech recordings are up to 30 seconds while two people conversations are of up to one hour long. The recordings are transcribed and segmented by speaker, noise, music, and overlapping speech.

Automatic speech recognition (ASR) is also known as speech-to-text and voice recognition.

What use cases is the data for?

The speech recognition datasets are perfect for:
- Building a speech recognition AI.
- Building a speaker recognition AI.
- Speech recognition solutions for call centers.

Dataset license

Our data licenses agreement covers commercial use, and the datasets can be reused for multiple cases. However, they are not for reselling.


Speech recognition sample collected through our service.

How is data relevant to speech recognition?

Read more about speech recognition here.

Quality guarantee

We are confident in our data, and all customers can review a sample batch of data before buying. Additionally, we offer a quality guarantee. If you wish to review more samples before buy, state so when filling in the order form.

Swedish speech data starting from

218€ / hour
Order now



16 – 44 kHz
Classified by noise level
Depends on case, up to 1 hour long.
Verbatim and/or read from sentences


16 – 85 years
Female 40%, Male 60%
Grouped by native and non-native Swedish speakers
Grouped by country of origin and region within the country
“Partnering with StageZero has been vital in providing us with high-quality utterance corpora for training our proprietary language and semantic models.”
Dr. Christoph Neumann
CTO at German Autolabs

Custom training data collection for speech and NLP.

Didn’t find the speech dataset you need or your industry in our marketplace? Get in touch with us, so we can use our global network to source the training or testing data that fits your needs.
Palkkatilanportti 1, 4th floor, 00240 Helsinki, Finland
©2022 StageZero Technologies
linkedin facebook pinterest youtube rss twitter instagram facebook-blank rss-blank linkedin-blank pinterest youtube twitter instagram