AI/ML Engineer

Remote · Full-timeApply by Nov 1, 2026

Role summary

You'll work on the conversation quality of Wabel's AI agents themselves — prompt design, evaluation, and the infrastructure around running and monitoring LLM-based conversations at scale across voice and text.

What you'll do

  • Design and iterate on prompts and conversation flows for different industries and use cases.
  • Build evaluation pipelines to catch quality regressions before they reach customers.
  • Work with customer success to diagnose and fix real conversation failures.
  • Stay current on the LLM/voice-AI landscape and bring in what's actually useful.

What we're looking for

  • Experience building production systems on top of LLM APIs (OpenAI, Anthropic, or similar).
  • Strong Python, comfortable writing evaluation/testing infrastructure, not just prompts.
  • Bilingual Arabic/English a strong plus, given the agents primarily serve Arabic-speaking customers.
  • Comfortable with ambiguity — this space moves fast and best practices are still forming.

Apply for this role

Tell us a bit about yourself and attach your CV — we'll get back to you if it looks like a fit.

* Required field

PDF, DOC, or DOCX — up to 5MB.

Let's tailor it to your business

We'll set everything up for a smooth launch — without the complexity.

Prefer to talk? Talk to Sales
No credit card · No technical setup