Arabic Speech Studio

Text-to-Speech & Speech-to-Text

0 TTS Models 0 ASR Models

Playground

Generate speech from text using any model

STT / ASR

Transcribe Arabic speech from an uploaded audio file

ASR

Drop a WAV/MP3 file here
or click to browse

Benchmark

Run all models on the same text and compare

Voice Cloning

Upload a reference voice (5–30 s) and clone it

🎤

Drop a WAV/MP3 file here
or click to browse

Tips for best results

  • Use 5–20 seconds of clean audio
  • Single speaker, minimal background noise
  • Providing a transcript improves accuracy
  • F5-TTS, SILMA, and Coqui XTTS-v2 give the strongest cloning options

Leaderboard

User ratings from this session

🇸🇦 Arabic Models

ModelRatingsAvg RatingAvg Gen (s)Avg RTF

🇬🇧 English Models

ModelRatingsAvg RatingAvg Gen (s)Avg RTF