Skip to content
Bayanat Voice

Arabic speech data for recognition and conversation.

Find audio for the way your users actually speak: on a phone, in a service call, or in an unscripted conversation. Browse Arabic and Saudi collections by recording type and training task.

Find your speech dataset

Choose a collection to see its counts, intended use, and licensing details.

Call-center conversations and scripted monologues make up the 10,000-hour shared catalog; ask us for each product’s allocation. The two Saudi 1,000-hour collections are separate products. Other collections may overlap, so check the release details before combining them.

Need more than the existing collection?

Our team can add data or annotation for your particular model. These additions are agreed separately.

Speech recognition

Scripted and spontaneous audio are listed separately, so you can choose the recording conditions and language coverage you need.

Conversation

Compare call-center, Saudi spontaneous, Saudi conversational, and full-duplex collections. Each has its own specifications and hour count.

Additional annotations

Request verbatim and normalized transcripts, speaker segments, timestamps, overlap, language switching, entities, or intent. Included fields vary by collection.

Voice-agent behavior

Add labels for corrections, clarification, confirmations, escalation, and resolution when you need to train beyond transcription.

How we put the data together

Choose a collection

Review source, dialect, speaker mix, and recording conditions against the conversations your product will handle.

Check the sample

Ask for audio, transcripts, and a delivery schema. Custom deliveries can include original recordings and separately labeled audio derivatives.

Set the terms

Confirm your intended use, required labels, and splits. Real interactions, role-play, and synthetic audio remain distinguishable.

A correction in the middle of a call

A caller changes a building number while the agent is speaking. The useful test is whether the agent hears the correction and saves the right address.

Common questions

Are the two 1,000-hour Saudi speech products the same?

No. Saudi Spontaneous Speech and Saudi Conversational Speech are separate collections, each with 1,000 hours. The conversational product includes human transcripts, timestamps, speaker attributes, and a fixed delivery schema.

Does the 10,000-hour catalog include 10,000 hours of each recording type?

No. That is the combined total for Arabic Call-Center Conversations and Arabic Scripted Monologues. They are separate products; request the hour allocation and speaker mix for each.

Can I use recognition data to create a synthetic voice?

That depends on the collection’s license and speaker permissions. Synthetic voice creation needs explicit terms. Our Studio category covers expressive recordings and custom voice production.