AI Engineer

Job Summary

Build and operate production AI and LLM systems on cloud infrastructure, leading a team to deliver domain-specific AI applications with safety guardrails and measurable performance metrics.

Responsibilities

  • Build, fine-tune, and evaluate LLM systems for domain-specific tasks using QLoRA/PEFT on open-weight models such as Llama-3 and Mistral.
  • Design reproducible evaluation harnesses and A/B test frameworks with tracked metrics including task success rate, safety rate, and latency distributions.
  • Architect multi-agent and RAG systems using LangGraph, FastAPI, and vector databases from prototype through production.
  • Implement safety guardrails including input/output validation, allowlist/denylist policies, and controls to reduce invalid or high-risk model actions.
  • Design and operate cloud infrastructure and MLOps workspaces on Kubernetes and containerized runtimes.
  • Build CI/CD pipelines and GitOps-based release promotion across development, test, and production environments.
  • Lead and mentor a cloud/AI operations team; define monitoring, incident response, and release governance practices.
  • Standardize SDLC practices including branching strategy, PR governance, and release management to improve delivery metrics.

Must haves

  • Bachelor’s degree in Software Engineering, Computer Science, or a related field.
  • 6–8+ years in software, DevOps, or platform engineering, including at least 2 years in applied AI or ML engineering.
  • Proven delivery of production AI/LLM systems, not research or notebook-stage work.
  • Strong Python; comfortable with Bash and YAML.
  • Deep hands-on experience with Kubernetes, Docker/Podman, and Terraform.
  • Production experience with at least one major cloud platform (Azure preferred; OCI or GCP acceptable).

Nice to haves

  • Master’s degree in Applied AI, Machine Learning, or a related discipline.
  • Fine-tuning experience with QLoRA/LoRA on GPU clusters; PyTorch and Transformers.
  • Vector database experience with Milvus, Pinecone, or Weaviate and RAG retrieval design.
  • Arabic and English professional proficiency.

We refresh listings regularly, but some roles close early on the source platform.

Country: Saudi Arabia
City: Riyadh
Job Category: AI/ML Engineering
Job Type: Full Time
Company Name: Saudi Azm عزم السعودية
Seniority level: Mid-Senior level
Sorry! This job has expired.
Scroll to Top