A passionate AI Evaluator, Prompt Engineer, and No-Code Platform Builder based in Riyadh, Saudi Arabia. I specialize in evaluating Large Language Model outputs, detecting hallucinations, and conducting adversarial testing (Red Teaming) — primarily in Arabic.
| ✅ | OpenAI Cookbook Contributor | PR #3049 merged |
| 🔴 | LLM Red Teamer | Adversarial Arabic prompts |
| ⚡ | No-Code AI Builder | API-integrated platforms |
| 🎵 | Educational Innovation | Arabic prosody × music timing |
| Category | Skills |
|---|---|
| 🎯 AI Evaluation | Hallucination Detection · Consistency Testing · LLM Benchmarking |
| 🔴 Security | Red Teaming · Adversarial Prompting · Jailbreak Analysis |
| ✍️ Engineering | Prompt Optimization · API Integration · Technical Writing (AR/EN) |
| 🤖 Machine Learning | Scikit-learn · Pipelines · Feature Engineering · Ensemble Methods |
| 🌐 Tools | Python · Google Colab · Git/GitHub · No-Code Platforms |
| Project | Description | Status |
|---|---|---|
| 🌊 AI Wave 404 | Systematic evaluation of LLM outputs in Arabic | 🔵 Active |
| 🎵 Arabic Prosody × Musical Timing | Interactive educational platform (No-Code) | ✅ Completed |
| 📊 Binary Classification Pipeline | End-to-end ML pipeline with ensemble methods (100% accuracy) | ✅ Completed |
| Role | Focus Area |
|---|---|
| 🧪 LLM Evaluator | Quality assessment & benchmarking |
| ✍️ Prompt Engineer | Optimization & system design |
| 🔴 Red Teamer | Security testing & vulnerability discovery |
| ⚡ No-Code AI Solutions | Rapid prototyping & deployment |
| 🤖 ML Engineer | Pipeline development & model optimization |

