AI Engineer
TLDR
Build and operate LLM models across GCP and sovereign cloud environments, covering Arabic NLP, GPU inference, monitoring, and compliance.
Looking for an AI Engineer to Develop, deploy, and operate AI/LLM models across Clinets dual environment — GCP for public-cloud workloads, Humain sovereign cloud for classified data.
Requirements
Build and fine-tune LLM/ML models for Arabic NLP, document classification, vision/OCR, and AIOps use cases.
Run pre-deployment evaluation
Accuracy baselines, regression and safety testing; evidence to justify GPU allocation.
Optimize inference — quantization, batching, context sizing — against measured usage.
Deploy on Humain GPUaaS: Kubernetes, GPU partitioning on B300 nodes, quotas, RBAC.
Build equivalent workloads on GCP (Vertex AI, GKE) with classification-based routing.
Own serving stack (vLLM/TGI), model versioning, CI/CD, and monitoring for latency, tokens, GPU utilization, and drift.
Ensuring developed AI Models Complying with ZATCA data sovereignty and SDAIA requirements (AI Ethics, GenAI Guidelines, PDPL).
Benefits
5 years ML/AI engineering, in production LLM deployment with knowledge in
Python, PyTorch, Hugging Face
Kubernetes in production; GPU-served inference
GCP Vertex AI or any equivellent cloud
InnovationTeam is a technology company at the forefront of the telecommunications industry, specializing in cloud, AI, and software solutions. We cater to diverse markets, delivering innovative products that empower businesses to thrive. Our mission is to create an ecosystem that enables motivated individuals to build rewarding careers in technology sales.
- Founded
- Founded 2009
- Employees
- 201-500 employees
- Industry
- information technology and services