الملفات
ghaymah-genai-exam/q2-arabic-tts-stt/arabic-tts-stt-proposal.md

40 أسطر
1.1 KiB
Markdown
خام الرابط الدائم اللوم التاريخ

هذا الملف يحتوي على أحرف Unicode غامضة

هذا الملف يحتوي على أحرف Unicode قد تُخلط مع أحرف أخرى. إذا كنت تعتقد أن هذا مقصود، يمكنك تجاهل هذا التحذير بأمان. استخدم زر الهروب للكشف عنها.

# Q2 — Propose an Arabic TTS or STT Model + Dataset
**Budget: ~6 minutes** · Weight: 15%
## Choose ONE option
- **Option A — TTS.** Propose a text-to-speech model to fine-tune for Arabic.
- **Option B — STT.** Propose a speech-to-text model to fine-tune for Arabic.
## The task
Propose **one model** and **one dataset**, with a short justification. No training, no running code — just a well-reasoned pick. (Short code/CLI snippets are allowed but not required.)
Cover:
1. **Model** — name it, and say why: architecture, license, multilingual/Arabic support.
2. **Dataset** — name it, and say why: size, dialect (MSA vs. Egyptian vs. Gulf vs. Maghrebi), license, where to get it.
3. **Why this pairing works for Arabic** — 12 sentences (e.g. diacritization handling, dialect coverage).
4. **One key risk** — 1 sentence.
## ANSWER — Arabic TTS/STT Proposal
**Option chosen:** A / B (delete one)
### Proposed model
> _write here_
### Proposed dataset
> _write here_
### Why this pairing works for Arabic
> _write here_
### One key risk
> _write here_