Skip to main content
Talk to Sales

AI Platform

Bring your containerized jobs. We run them — on the tier that fits.

The AI Platform is a ready-to-deploy environment where you submit your own containerized jobs without managing servers. Each job runs on the tier you choose — a heavily discounted flexible tier for interruptible work, or the production tier — all orchestrated with Kubernetes.

Containers
bring your own, package once
Two tiers
flexible and production
Kubernetes
orchestrated scheduling and scaling

How tiering works

Two ways to run the same job

You decide what matters more for each job — cost or certainty. The platform handles the rest.

Flexible tier

Interruptible, heavily discounted

Jobs that can pause and resume — batch processing, experimentation, pre-training runs — run on spare capacity at a deep discount. They get scheduled when capacity is available and are safely paused and resumed when it is not.

  • Heavily discounted compute
  • Pause and resume as capacity allows
  • Ideal for batch, research, and pre-training

Production tier

Top-tier performance, guaranteed priority

For jobs that must finish on time — production inference, client-facing workloads — the production tier reserves capacity and gives your jobs scheduling priority, with stable performance guaranteed.

  • Priority scheduling with reserved capacity
  • Consistent, top-tier performance
  • Ideal for production inference and live workloads

How it works

From container to running job in three steps

No servers to provision and no clusters to manage. Your team keeps building software the way it already does — in containers.

01

Package your job

Build your workload as a standard container image — the same format your engineers already use every day.

02

Submit it

Send the job to the platform and choose the tier. No capacity planning, no infrastructure tickets.

03

We schedule and scale

Kubernetes-based orchestration schedules the job across GPU capacity, autoscales as needed, and recovers it automatically if anything fails.

Billing model

Tiered usage — you pay for what you run, on the tier you run it

Usage is metered and billed according to the tier you choose. Run experimental jobs on the flexible tier when cost matters most; run production jobs on the production tier when it does. The financing dashboard tracks it all in real time.

Metered by tier

Each job is billed against the tier it ran on.

Pay for what you run

No idle reservations — you pay for the compute your jobs actually use.

Full visibility

Real-time usage and spend tracking, with budget limits and downloadable invoices.

Use cases

Built for teams that ship AI

One platform that scales from experiments to production without re-platforming.

Batch training and fine-tuning

Long-running jobs on the flexible tier, scheduled around available capacity at a deep discount.

Production inference

Client-facing workloads on the production tier, with priority scheduling and consistent performance.

Experiments and CI for ML

Rapid iteration and automated model evaluation — spin jobs up and down without managing infrastructure.

Run your first job on the platform

Tell us about your workloads, the container images you run, and your performance requirements. A senior member of our team will respond within one business day.