Looking for an AI Engineer to Develop, deploy, and operate AI/LLM models across Clinets dual environment — GCP for public-cloud workloads, Humain sovereign cloud for classified data.
Requirements
Build and fine-tune LLM/ML models for Arabic NLP, document classification, vision/OCR, and AIOps use cases.
Run pre-deployment evaluation
Accuracy baselines, regression and safety testing; evidence to justify GPU allocation.
Optimize inference — quantization, batching, context sizing — against measured usage.
Deploy on Humain GPUaaS: Kubernetes, GPU partitioning on B300 nodes, quotas, RBAC.
Build equivalent workloads on GCP (Vertex AI, GKE) with classification-based routing.
Own serving stack (vLLM/TGI), model versioning, CI/CD, and monitoring for latency, tokens, GPU utilization, and drift.
Ensuring developed AI Models Complying with ZATCA data sovereignty and SDAIA requirements (AI Ethics, GenAI Guidelines, PDPL).
Benefits
5 years ML/AI engineering, in production LLM deployment with knowledge in
Python, PyTorch, Hugging Face
Kubernetes in production; GPU-served inference
GCP Vertex AI or any equivellent cloud
Skills Required
- Build and fine-tune LLM/ML models for Arabic NLP, document classification, vision/OCR, and AIOps use cases.
- Run pre-deployment evaluation: accuracy baselines, regression and safety testing; provide evidence to justify GPU allocation.
- Optimize inference using quantization, batching, and context sizing against measured usage.
- Deploy on Humain GPUaaS: Kubernetes, GPU partitioning on B300 nodes, quotas, and RBAC.
- Build equivalent workloads on GCP (Vertex AI, GKE) with classification-based routing between environments.
- Own serving stack (vLLM/TGI), model versioning, CI/CD, and monitoring for latency, tokens, GPU utilization, and drift.
- Ensure AI model compliance with ZATCA data sovereignty, SDAIA requirements, AI Ethics, GenAI Guidelines, and PDPL.
- 5 years ML/AI engineering experience with production LLM deployment.
- Proficiency in Python, PyTorch, and Hugging Face.
- Production Kubernetes experience and GPU-served inference knowledge.
- Experience with GCP Vertex AI or equivalent cloud ML platforms.
What We Do
Innovation Team is an IT consulting company that provides a comprehensive range of specialized professional solutions and services to businesses. Headquartered in Toronto, Canada, and branches serving other regions in the world, Innovation Team seeks to assist businesses operating in various industries to achieve their business objectives and to perform their day-to-day operations as competently and as efficiently possible. Our solutions and services are delivered at the hands of some of the most dedicated professionals in the field. In providing our scope of offerings, we work closely with vendors/manufacturers, experienced consulting firms, and system integrators; ensuring all our clientele receive optimum and timely services, consistently.






