Guide hub

Guides for AI workload execution, GPU cost, and LLM deployment

Start here if you are working through workload routing, GPU cost, failover behavior, model fit, and the practical tradeoffs of running AI workloads across fragmented capacity.

Estimate costBrowse model pages
Best for
Practical questions
Use these guides when you need operational answers, not marketing copy.
Coverage
Cost, fit, failover
The library focuses on the deployment questions teams hit most often.
Common starting point
Inference first
Start with inference if your application needs a trained model to answer requests.

What you will find

Start with the practical questions teams ask first

These guides focus on the questions that come up once a team moves from experimenting with models to shipping them reliably. That means cost, fit, fallback behavior, and how much provider-specific logic you really want to own.

Use the guides to understand the problem first, then branch into model-specific pages or pricing when you want a more concrete route.