Senior AI DevOps / LLMOps
Analyse Teletravail.ma
Éligibilité Maroc : Maroc accepté
Cette offre accepte les candidats basés au Maroc.
Pourquoi ? L'annonce est ouverte au monde entier, sans restriction de pays ni exclusion du Maroc.
Déduit des pays indiqués dans l'annonce d'origine. Vérifiez toujours les conditions avant de postuler.
Faits de l'annonce d'origine (Himalayas)
- Entreprise
- TechBiz Global
- Mode de travail
- 100 % à distance
- Localisation autorisée
- Worldwide
- Contrat
- Temps plein
- Niveau
- Senior
- Métier
- Développement
- Publiée le
- 22 septembre 2026
- Expire le
- 21 novembre 2026
- Source
- Himalayas
- Dernière vérification
Description de l'offre
Texte d'origine publié par l'employeur, non modifié.
At TechBiz Global, we are providing recruitment service to our TOP clients from our portfolio. We are currently seeking an Senior AI DevOps / LLMOpsspecialist to join one of our clients' teams. If you're looking for an exciting opportunity to grow in a innovative environment, this could be the perfect fit for you.
Key Responsibilities
Automation of Build-to-Production
dataset versioning, and application code.
- Develop specialized workflows for PromptOps, ensuring that system prompts are
version-controlled, tested for regressions, and deployed with the same rigor as traditional
code.
- Automate the deployment of Agentic workflows, managing the complexities of stateful
AI interactions and multi-agent handoffs.
2. AI Infrastructure as Code (IaC)
- Provision and manage high-performance compute environments (GPU clusters, TPU
pods) using Terraform, Pulumi, or Ansible.
- Define and enforce Policy-as-Code for AI endpoints to ensure compliance with security,
cost-usage limits, and data residency requirements.
- Maintain a consistent environment across Hybrid Infrastructure, ensuring seamless
parity between On-Premises development and Cloud production.
3. Safe Experimentation & Controlled Releases
- Architect Progressive Delivery strategies for AI, including Canary releases, Blue-Green
deployments, and Shadowing (where new models run in parallel with production to
compare outputs).
- Build “Evaluation-in-the-Loop” gates within the pipeline to automatically test for bias,
hallucination, and performance degradation before a release.
- Implement A/B testing frameworks specifically designed for LLM outputs and agentic
behavior.
4. Monitoring & Observability
- Establish deep observability into Inference Endpoints, tracking metrics like tokens-per-
second, latency, and drift in model accuracy.
- Integrate feedback loops that capture production “edge cases” to feed back into the
training and fine-tuning pipelines.
Requirements
Must-Have Technical Skills:
- Orchestration: Advanced Kubernetes (K8s) skills, specifically with KubeFlow, Ray, or
NVIDIA Triton.
- CI/CD & IaC: Expertise in GitHub Actions/GitLab CI, and Terraform or Pulumi.
- AI Tooling: Experience with Weights & Biases, MLflow, LangSmith, or Arize
Phoenix.
- Hardware: Understanding of GPU virtualization, CUDA drivers, and on-premises
hardware management.
- Security: Familiarity with Open Policy Agent (OPA) and secret management (Vault).
Experience:
- 10+ years in DevOps, SRE, or Cloud Engineering.
- 2+ years of hands-on experience in MLOps or LLMOps, specifically moving LLMs
from notebook to production.
- Proven experience managing Hybrid Cloud environments (e.g., AWS/Azure + Private
Data Center).
Highlights
- full time and remote job
- fluent English is needed
Originally posted on Himalayas
TechBiz Global
La source ne fournit pas de présentation de l'entreprise.
Source
Himalayas — Voir l'annonce originale
Toujours publiée sur Himalayas — vérifiée il y a 1 h.
Teletravail.ma référence cette annonce et n'est pas l'employeur.
