Fine-tuning & alignment
When a prompt is not enough, we fine-tune. Supervised runs on your labelled data, preference optimisation for tone and refusal behaviour, and a held-out Arabic evaluation set so gains are real, not vibes.
- SFT & preference (DPO/ORPO)
- Arabic + English eval sets
- Model & weights handed back
from AED 60,000 · 4–6 weeks
