Speech recognition
Scripted and spontaneous audio are listed separately, so you can choose the recording conditions and language coverage you need.
Find audio for the way your users actually speak: on a phone, in a service call, or in an unscripted conversation. Browse Arabic and Saudi collections by recording type and training task.
Choose a collection to see its counts, intended use, and licensing details.
For speech and contact-center teams that need conversation audio rather than individual scripted recordings.
View dataset Part of our 10,000-hour speech catalogFor ASR teams that need scripted recordings with a clear source category for acoustic modeling and recognition experiments.
View dataset 500 hoursFor teams training mobile speech recognition or assistants that need to understand Arabic recorded on a phone.
View dataset 1,000 hoursFor ASR and spoken-language teams that need Saudi speech beyond read scripts.
View dataset 1,000 hoursFor teams that need ready-to-license Saudi conversation data with human transcripts and a defined delivery format.
View dataset 350 hoursFor voice-agent teams working on customer-service interaction and evaluating the labels needed for turn-taking, corrections, and task completion.
View dataset Ask about available volumeFor teams developing expressive Arabic voices that need professional recordings and explicit terms for their model use.
View dataset 2,000 prompts · 40 speakers · 2 hoursFor teams that need a small, clearly defined studio collection to test a speech interaction or explore a voice prototype.
View datasetCall-center conversations and scripted monologues make up the 10,000-hour shared catalog; ask us for each product’s allocation. The two Saudi 1,000-hour collections are separate products. Other collections may overlap, so check the release details before combining them.
Our team can add data or annotation for your particular model. These additions are agreed separately.
Scripted and spontaneous audio are listed separately, so you can choose the recording conditions and language coverage you need.
Compare call-center, Saudi spontaneous, Saudi conversational, and full-duplex collections. Each has its own specifications and hour count.
Request verbatim and normalized transcripts, speaker segments, timestamps, overlap, language switching, entities, or intent. Included fields vary by collection.
Add labels for corrections, clarification, confirmations, escalation, and resolution when you need to train beyond transcription.
Review source, dialect, speaker mix, and recording conditions against the conversations your product will handle.
Ask for audio, transcripts, and a delivery schema. Custom deliveries can include original recordings and separately labeled audio derivatives.
Confirm your intended use, required labels, and splits. Real interactions, role-play, and synthetic audio remain distinguishable.
A caller changes a building number while the agent is speaking. The useful test is whether the agent hears the correction and saves the right address.
No. Saudi Spontaneous Speech and Saudi Conversational Speech are separate collections, each with 1,000 hours. The conversational product includes human transcripts, timestamps, speaker attributes, and a fixed delivery schema.
No. That is the combined total for Arabic Call-Center Conversations and Arabic Scripted Monologues. They are separate products; request the hour allocation and speaker mix for each.
That depends on the collection’s license and speaker permissions. Synthetic voice creation needs explicit terms. Our Studio category covers expressive recordings and custom voice production.