Invent a Better Everyday | Abu Dhabi, UAE | G42

Make Enquiry

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Scaling an Agentic AI Platform 3.5x

Scaling an Agentic AI Platform 3.5x

How Core42's sovereign AI cloud gave AIREV the elastic, resilient infrastructure to scale OnDemand's 300+ production agents.
 

 

THE CHALLENGE
 

OnDemand coordinates more than 300 AI agents across 16 microservices. As the platform expanded, AIREV needed GPU infrastructure that could grow rapidly, maintain enterprise-grade uptime and place different inference workloads on the most suitable hardware. Enterprise deployments also required UAE data sovereignty and regional governance alignment. Fixed capacity, provisioning delays or dependence on one accelerator family would constrain performance, economics and speed to market.

 

 

THE CORE42 SOLUTION
 

Core42 delivered a fully managed, sovereign AI cloud that combines elastic capacity, multi-accelerator choice and production-grade operations:

 

→  Elastic GPU scaling — capacity expands on demand without provisioning delays.

→  Mixed inferencing — NVIDIA, AMD and Cerebras match workload price-performance needs.

→  Managed AI operations — Kubernetes, Slurm and built-in observability unify the environment.

→  Multi-zone sovereignty — 99.95% availability with UAE data residency and governance.

PLATFORM CAPABILITIES

 

CAPABILITIES WHAT IT DOES
Elastic GPU Infrastructure Scales on demand without provisioning delays, giving a growing agentic platform capacity when production workloads increase.
Multi-Accelerator Inference Offers NVIDIA, AMD and Cerebras options so workloads can be aligned with different performance-cost profiles.
Unified AI Operations Integrates Kubernetes and Slurm orchestration with built-in observability in a fully managed operating environment.
Multi-Zone Sovereign Architecture
Supports a 99.95% availability SLA while aligning enterprise deployments with UAE data residency and governance requirements.

 

DEEP DIVE DEEP DIVE DEEP DIVE DEEP DIVE

OnDemand is AIREV's agentic AI operating system for building, deploying and managing AI applications. Behind the user experience, more than 300 AI agents operate across 16 microservices, creating a fast-changing infrastructure profile. Agent workloads rise unevenly, inference economics vary by model and accelerator, and enterprise customers expect high availability alongside regional control of data. AIREV therefore needed a production environment that could scale quickly, absorb changing demand and support new deployment models without locking the platform to one hardware architecture.

To achieve this, Core42 delivered a fully managed, sovereign AI cloud rather than a fixed GPU cluster. Elastic capacity expands without provisioning delays, while accelerator choice across NVIDIA, AMD and Cerebras allows each inference workload to be aligned with an appropriate performance-cost profile. Kubernetes and Slurm provide common orchestration, and built-in observability gives AIREV one operating view across the environment. A multi-zone architecture adds resilience, while UAE-based deployment supports data residency and governance requirements for regional enterprise customers.

With this foundation, AIREV scaled compute 3.5x, with the environment carrying a 99.95% availability SLA and supportingmixed inferencing across all three accelerator families. OnDemand Enterprise is also expanding through software-only, self-hosted and cloud-hosted distribution models across the GCC and Africa. Core42's infrastructure gives AIREV a flexible production base for that next phase: capacity can grow with adoption, workloads can move to the right hardware, and enterprise deployments can maintain the resilience, governance and sovereign readiness required in each market.

OUTCOMES AT A GLANCE

 

3.5x

Compute scale achieved

US $1.225M

Reported 2025 compute spend

99.95%

Availability SLA

300+

AI agents in production

16

Microservices supporting OnDemand

3

Accelerator families used for inference

 

Related News

For better web experience, please use the website in portrait mode