AI inference platform offering model APIs, dedicated endpoints, and containers for fast and affordable LLM inference.
Active plans
7
Total changes
7
Last change
Jun 5, 2026
Price trend
Changes in 12 months
7
Cadence
Multiple changes per day
Direction
New plan recently added
Category position
Cheaper than 97% of AI & ML tools (2268 compared)
From $0.10/1M tokens (Llama-3.1-8B-Instruct)
$2.9/hour (A100 80GB)
$3.9/hour (H100 80GB)
$4.5/hour (H200 141GB)
$8.9/hour (B200 180GB)
Last 7 changes
| Date | Plan | Change | Old | New | Δ% |
|---|---|---|---|---|---|
| Jun 5, 2026 | Enterprise | new plan | — | — | — |
| Jun 5, 2026 | Container | new plan | — | — | — |
| Jun 5, 2026 | Dedicated Endpoints B200 | new plan | — | $8.9 | — |
| Jun 5, 2026 | Dedicated Endpoints H200 | new plan | — | $4.5 | — |
| Jun 5, 2026 | Dedicated Endpoints H100 | new plan | — | $3.9 | — |
| Jun 5, 2026 | Dedicated Endpoints A100 | new plan | — | $2.9 | — |
| Jun 5, 2026 | Model APIs (Pay-per-token) | new plan | — | $0.1 | — |
No confirmed competitors yet. These are AI & ML products at a similar price, matched automatically — not verified alternatives.