AI inference platform offering model APIs, dedicated endpoints, and containers for fast and affordable LLM inference.
Active plans
7
Total changes
7
Last change
Jun 5, 2026
Price trend
Changes in 12 months
7
Cadence
Multiple changes per day
Direction
New plan recently added
Category position
—
Not enough peers tracked yet
From $0.10/1M tokens (Llama-3.1-8B-Instruct)
Last 7 changes
| Date | Plan | Change | Old | New | Δ% |
|---|---|---|---|---|---|
| Jun 5, 2026 | Enterprise | new plan | — | — | — |
| Jun 5, 2026 | Container | new plan |
Be the first to flag a competitor.
$2.9/hour (A100 80GB)
$3.9/hour (H100 80GB)
$4.5/hour (H200 141GB)
$8.9/hour (B200 180GB)
| — |
| — |
| — |
| Jun 5, 2026 | Dedicated Endpoints B200 | new plan | — | $8.9 | — |
| Jun 5, 2026 | Dedicated Endpoints H200 | new plan | — | $4.5 | — |
| Jun 5, 2026 | Dedicated Endpoints H100 | new plan | — | $3.9 | — |
| Jun 5, 2026 | Dedicated Endpoints A100 | new plan | — | $2.9 | — |
| Jun 5, 2026 | Model APIs (Pay-per-token) | new plan | — | $0.1 | — |
© 2026 PriceTrack. All rights reserved.
Last deploy main@a1b2c3d