PriceTrack
ProductsChangesBenchmarksComparePricing
Products/AI & ML/FriendliAI

FriendliAI

Automated

AI inference platform offering model APIs, dedicated endpoints, and containers for fast and affordable LLM inference.

Visit websitePricing page

Active plans

7

Total changes

7

Last change

Jun 5, 2026

Price trend

no history

What we know about FriendliAI’s pricing

Volatile

Changes in 12 months

7

Cadence

Multiple changes per day

Direction

New plan recently added

Category position

Cheaper than 97% of AI & ML tools (2268 compared)

Current pricing

Model APIs (Pay-per-token)

usage based
$0.1

From $0.10/1M tokens (Llama-3.1-8B-Instruct)

  • Pay per token
  • Frontier model inference
  • GLM-5.1: $1.4 input/$4.4 output
  • Llama-3.3-70B: $0.6
  • DeepSeek-V3.2: $0.5 input/$1.5 output
  • Whisper-large-v3: $0.0015/audio minute

Dedicated Endpoints A100

usage based
$2.9

$2.9/hour (A100 80GB)

  • Pay per second
  • No start-up charges
  • On-demand deployment

Dedicated Endpoints H100

usage based
$3.9

$3.9/hour (H100 80GB)

  • Pay per second

Dedicated Endpoints H200

usage based
$4.5

$4.5/hour (H200 141GB)

  • Pay per second

Dedicated Endpoints B200

usage based
$8.9

$8.9/hour (B200 180GB)

  • Pay per second

Container

custom
Custom
  • Run inference with full control
  • Inference in your environment

Enterprise

custom
Custom
  • Custom Model APIs rate limits
  • Priority GPU access
  • Reserved GPU capacity
  • VPC deployments
  • On-prem deployment options
  • Dedicated support
  • Named Customer Success

Price history

Last 7 changes

DatePlanChangeOldNewΔ%
Jun 5, 2026Enterprisenew plan———
Jun 5, 2026Containernew plan———
Jun 5, 2026Dedicated Endpoints B200new plan—$8.9—
Jun 5, 2026Dedicated Endpoints H200new plan—$4.5—
Jun 5, 2026Dedicated Endpoints H100new plan—$3.9—
Jun 5, 2026Dedicated Endpoints A100new plan—$2.9—
Jun 5, 2026Model APIs (Pay-per-token)new plan—$0.1—

Similar to FriendliAI.

No confirmed competitors yet. These are AI & ML products at a similar price, matched automatically — not verified alternatives.

Novita AI

From Usage-based per model (e.g. DeepSeek V4 Pro: $1.60/1M in)

AI inference API with 100+ models including LLM, image, audio, and video generation with serverless and dedicated endpoints.

Cerebras Inference

From $50/month

Ultra-fast AI inference API powered by Cerebras hardware delivering 20x faster speeds than GPU-based providers for open-source and production models.

Fireworks AI

From $0.008 per 1M input tokens

Fast inference API for open-source generative AI models

AI21 Labs

From $0.2/1M tokens

Foundation model and AI API platform offering Jamba models for enterprise generative AI applications.

ModelsLab

From $21/month

AI API platform for image, video, audio and LLM model inference, replacing multiple API providers.

Cerebras Inference

From $10

Ultra-fast AI inference API powered by Cerebras Wafer-Scale processors, delivering the world's fastest LLM inference speeds.

PriceTrack

The source of truth for SaaS pricing. Historical data, instant alerts, a developer API.

Product

  • Browse
  • Compare
  • Trends
  • Pricing

Developers

  • API Docs
  • MCP server
  • Dashboard
  • Webhooks

Founders

  • Claim your product
  • Founder dashboard

Company

  • Changelog
  • About
  • Privacy
  • Terms
  • Refunds

© 2026 PriceTrack. All rights reserved.

support@pricetrack.dev