All Anyscale Plans & Pricing

Plan Monthly Annual Best For
View all features by plan (compare side-by-side)

Pay As You Go

  • $100 free credit on sign-up
  • Hosted managed infrastructure
  • CPU and GPU compute (T4, L4, A10G, A100)
  • No monthly fixed fees — pay only for compute used
  • Business hours support with 5 case submissions
  • Limited cloud regions

Committed Contracts (BYOC)

  • Volume discounts that scale with usage
  • Bring Your Own Cloud — AWS, GCP, Azure, or on-prem
  • Deploy on VMs or Kubernetes
  • Data residency in your own VPC
  • Use existing GPU reservations
  • Invoice via Anyscale or cloud marketplace
  • Enterprise SLAs with 24x7 coverage
  • Unlimited support case submissions
Pricing Alerts

Track Anyscale pricing

Get an email when Anyscale's pricing changes — plus the weekly SaaS Price Watch: verified price changes and deals across 3,000+ products. One-click unsubscribe.

Compare Anyscale with alternativesAdjust seats, lock a tier, add up to 2 more products side-by-side. Shareable URL.
Quick Answer
Last verified:
Medium confidence

Anyscale offers usage-based pricing from $0.013–$4.96 per million tokens as of August 2026 and custom pricing for larger requirements. Plan: Pay As You Go (usage-based). Custom pricing is available on request. Usage cost depends on the selected model and generation volume.

Use the interactive pricing calculator to estimate your exact cost based on team size and requirements.

  • Free tier: No free tier available

Anyscale offers 2 pricing tiers: Pay As You Go, Committed Contracts (BYOC). Paid plans include Pay As You Go (usage-based). The Committed Contracts (BYOC) plan is enterprise ml teams with sustained gpu utilization needing volume discounts, data residency, and custom deployment.

Compared to other llm api providers software, Anyscale is positioned at the budget-friendly price point.

  • 7 documented hidden costs beyond list price

How much does Anyscale cost?

Anyscale offers usage-based pricing from $0.013–$4.96 per million tokens, with custom pricing available for larger requirements. Plans include Pay As You Go (usage-based), Committed Contracts (BYOC) (custom pricing).

Anyscale Pricing Overview

Anyscale offers usage-based pricing from $0.013–$4.96 per million tokens and custom pricing for larger requirements. The Pay As You Go plan is usage-based and is designed for teams getting started with ray-based ml workloads on fully managed hosted infrastructure. The Committed Contracts (BYOC) plan requires contacting sales for a custom quote and is designed for enterprise ml teams with sustained gpu utilization needing volume discounts, data residency, and custom deployment.

There are at least 7 documented hidden costs beyond Anyscale's list price, including implementation, training, and add-on fees.

This pricing was last verified in July 27, 2026 from 2 independent sources.

Anyscale is the commercial platform built on Ray, the distributed computing framework. It offers Anyscale Endpoints — a serverless LLM inference API with per-token pricing for open-source models — and managed Ray clusters for training and fine-tuning. Anyscale Endpoints are OpenAI-compatible, supporting Llama, Mistral, Mixtral, and other popular models. The platform is designed for teams that need both production-ready LLM APIs and the ability to run custom distributed workloads on Ray.

How Anyscale Pricing Compares

Compare Anyscale pricing against top alternatives in LLM API Providers.

Usage-Based Rates

Per-unit pricing for Anyscale API usage.

Pay As You Go

Model Unit Rate
cpu-only hour $0.013
nvidia-t4 hour $0.568
nvidia-l4 hour $0.954
nvidia-a10g hour $1.36
nvidia-a100 hour $4.96
  • Prices in Anyscale Credits (AC); 1 AC = $1 USD
  • NVIDIA H/B/GB GPU families (H100, B200, GB200) require contacting sales
  • Volume discounts available via committed contracts

Compare Anyscale vs Alternatives

Before committing to Anyscale, compare pricing with these 3 alternatives in the same category.

All Anyscale alternatives & migration guides

What Companies Actually Pay for Anyscale

Review scores
Third-party review aggregates, as of Jul 2026
Top pricing complaints
a noticeable learning curve, particularly for teams unfamiliar with Ray conceptsa lack of transparency in pricing, which makes cost planning challengingdifficulty during debugging

How Anyscale Pricing Compares

Software Starting Price Top Price
Anyscale Free $4.9591 per million tokens
Amazon Bedrock $0.07 per million tokens $75 per million tokens
Baidu ERNIE API $0.1 per million tokens $10 per million tokens
Cerebras Inference API $0.1 per million tokens $6 per million tokens
Claude API $0.1 per million tokens $75 per million tokens
Cloudflare Workers AI Free $4.881 per million tokens

7 Anyscale Hidden Costs Beyond the List Price

Beyond the listed price, Anyscale has at least 7 documented hidden costs that can significantly increase total cost of ownership.

Watch for 7 hidden costs
  • Compute Costs (GPUs) AC 0.0135/hr to AC 4.9591/hr
    high 1 source
    industry "Anyscale claims to offer up to 10x more cost-effective solutions for open-source LLMs compared to proprietary solutions for general workloads, and up to 6x cost savings for batch LLM inference in shared prefix scenarios compared to AWS Bedrock."
  • Developer Time / Learning Curve
    high 1 source
    industry "Anyscale claims to offer up to 10x more cost-effective solutions for open-source LLMs compared to proprietary solutions for general workloads, and up to 6x cost savings for batch LLM inference in shared prefix scenarios compared to AWS Bedrock."
  • Idle Resources
    medium 1 source
    industry "Idle Resources: While Anyscale offers features like auto-suspension for clusters to prevent paying for idle resources, managing this effectively is crucial for cost control."
  • Data Egress and Storage
    medium 1 source
    industry "Data Egress and Storage: Although not explicitly detailed for Anyscale, general LLM API pricing issues suggest that published rates might not capture the total cost of ownership, potentially excluding data egress, storage, or premium features that..."
  • Unpredictable Usage
    medium 1 source
    industry "Unpredictable Usage: For usage-based models, fluctuating workloads can make precise budgeting challenging, as costs depend heavily on real-time resource consumption."
  • Fine-tuning Fee $5 per run
    low 1 source
    industry "Fine-tuning and Inference: Anyscale charges a static fee of $5 per run for fine-tuning, regardless of the training data size."
  • LLM Inference Costs $1 per million tokens
    low 1 source
    industry "For LLM inference, Anyscale Endpoints are offered at $1 per million tokens for models like Llama-2 70B, and even less for other models."
Tip

Ask your Anyscale sales rep about these costs upfront. Getting them in writing before signing can save you from surprise charges later.

Full hidden costs breakdown →

Intelligence sourced from 1 independent sources
industry
Key claims include inline source attribution. Data verified against multiple independent sources. 10 source citations total.

How to Negotiate Anyscale Pricing

Anyscale contracts are negotiable. These 3 tactics are sourced from real buyer experiences and procurement specialists.

Negotiation Playbook 3 tactics
Negotiate Volume Discounts high success

For committed contracts, buyers can negotiate for volume discounts by committing to a certain level of spending.

https://g2.com
Optimize Resource Utilization high success

Anyscale's platform allows for efficient GPU utilization, auto-scaling, and the use of spot instances which can be 50-80% cheaper.

https://lawinsider.com
Push for Pricing Transparency medium success

Buyers should push for detailed breakdowns of all potential costs, including those beyond direct compute, as Anyscale's pricing structure can feel unclear.

https://g2.com

Full negotiation guide →

Anyscale Price History

Pricing changes CostBench has tracked for Anyscale, verified across 2 snapshots going back to 2026Q2.

  1. Anyscale tier renamed

    "Anyscale Endpoints" renamed to "Pay As You Go"

  2. Anyscale tier renamed

    "Managed Ray Clusters" renamed to "Committed Contracts (BYOC)"

Anyscale Pricing FAQ

01 How much do Anyscale Endpoints cost?

Anyscale Endpoints charge per token. Small models like Llama 3.1 8B and Mistral 7B cost $0.15 per million tokens (input and output same rate). Llama 3.1 70B costs $1.00/M tokens. Llama 3.1 405B costs $5.00/M tokens. New accounts get $10 in free credits.

02 What is Anyscale built on?

Anyscale is built on Ray, an open-source distributed computing framework developed at UC Berkeley. Anyscale provides the managed, enterprise version of Ray with production-grade SLAs, managed clusters, and hosted inference endpoints.

03 Does Anyscale have a free tier?

Anyscale gives new accounts $10 in free credits for Endpoints usage. There is no permanently free tier — after credits are used, standard per-token rates apply.

04 Anyscale vs Together AI: which should I use?

Together AI and Anyscale both offer OpenAI-compatible inference for open-source models. Together AI has broader model selection and slightly more competitive pricing for commodity models. Anyscale is better if you're already using Ray for training or need the Ray ecosystem integration.

Is this pricing incorrect? — we'll verify and update it.