Anyscale Pricing 2026
Complete pricing guide with plans, hidden costs, and cost analysis
Anyscale uses usage-based pricing from $0.013–$4.96 per million tokens and custom pricing for larger requirements.
Are you Anyscale? Claim this profile
Is Anyscale a fit?
Derived from this product's own pricing and contract record — not a review.
Best for
- Teams getting started with Ray-based ML workloads on fully managed hosted infrastructure
- Enterprise ML teams with sustained GPU utilization needing volume discounts, data residency, and custom deployment
Not a fit if
- The sticker price is your whole budget — Compute Costs (GPUs) (and 1 more).
All Anyscale Plans & Pricing
| Plan | Monthly | Annual | Best For |
|---|---|---|---|
| Pay As You Go freeCredit: $100 on sign-upsupport: Business hours only, 5 case submissions | Usage-based | Not published | Teams getting started with Ray-based ML workloads on fully managed hosted infrastructure |
| Verified pricing · last checked July 2026 · 2 sources
Get this price at Anyscale →
| |||
| What's included at Pay As You Go Best for: Teams getting started with Ray-based ML workloads on fully managed hosted infrastructure
Limits
| |||
| Committed Contracts (BYOC) | Contact Sales | Contact Sales | Enterprise ML teams with sustained GPU utilization needing volume discounts, data residency, and custom deployment |
| Verified pricing · last checked July 2026 · 2 sources
Get this price at Anyscale →
| |||
| What's included at Committed Contracts (BYOC) Best for: Enterprise ML teams with sustained GPU utilization needing volume discounts, data residency, and custom deployment
| |||
View all features by plan (compare side-by-side)
Pay As You Go
- $100 free credit on sign-up
- Hosted managed infrastructure
- CPU and GPU compute (T4, L4, A10G, A100)
- No monthly fixed fees — pay only for compute used
- Business hours support with 5 case submissions
- Limited cloud regions
Committed Contracts (BYOC)
- Volume discounts that scale with usage
- Bring Your Own Cloud — AWS, GCP, Azure, or on-prem
- Deploy on VMs or Kubernetes
- Data residency in your own VPC
- Use existing GPU reservations
- Invoice via Anyscale or cloud marketplace
- Enterprise SLAs with 24x7 coverage
- Unlimited support case submissions
Track Anyscale pricing
Get an email when Anyscale's pricing changes — plus the weekly SaaS Price Watch: verified price changes and deals across 3,000+ products. One-click unsubscribe.
You're on the list — first digest lands Tuesday.
Anyscale offers usage-based pricing from $0.013–$4.96 per million tokens as of August 2026 and custom pricing for larger requirements. Plan: Pay As You Go (usage-based). Custom pricing is available on request. Usage cost depends on the selected model and generation volume.
Use the interactive pricing calculator to estimate your exact cost based on team size and requirements.
- Free tier: No free tier available
Anyscale offers 2 pricing tiers: Pay As You Go, Committed Contracts (BYOC). Paid plans include Pay As You Go (usage-based). The Committed Contracts (BYOC) plan is enterprise ml teams with sustained gpu utilization needing volume discounts, data residency, and custom deployment.
Compared to other llm api providers software, Anyscale is positioned at the budget-friendly price point.
- 7 documented hidden costs beyond list price
How much does Anyscale cost?
Anyscale Pricing Overview
Anyscale offers usage-based pricing from $0.013–$4.96 per million tokens and custom pricing for larger requirements. The Pay As You Go plan is usage-based and is designed for teams getting started with ray-based ml workloads on fully managed hosted infrastructure. The Committed Contracts (BYOC) plan requires contacting sales for a custom quote and is designed for enterprise ml teams with sustained gpu utilization needing volume discounts, data residency, and custom deployment.
There are at least 7 documented hidden costs beyond Anyscale's list price, including implementation, training, and add-on fees.
This pricing was last verified in July 27, 2026 from 2 independent sources.
Anyscale is the commercial platform built on Ray, the distributed computing framework. It offers Anyscale Endpoints — a serverless LLM inference API with per-token pricing for open-source models — and managed Ray clusters for training and fine-tuning. Anyscale Endpoints are OpenAI-compatible, supporting Llama, Mistral, Mixtral, and other popular models. The platform is designed for teams that need both production-ready LLM APIs and the ability to run custom distributed workloads on Ray.
How Anyscale Pricing Compares
Compare Anyscale pricing against top alternatives in LLM API Providers.
Usage-Based Rates
Per-unit pricing for Anyscale API usage.
Pay As You Go
| Model | Unit | Rate |
|---|---|---|
| cpu-only | hour | $0.013 |
| nvidia-t4 | hour | $0.568 |
| nvidia-l4 | hour | $0.954 |
| nvidia-a10g | hour | $1.36 |
| nvidia-a100 | hour | $4.96 |
- Prices in Anyscale Credits (AC); 1 AC = $1 USD
- NVIDIA H/B/GB GPU families (H100, B200, GB200) require contacting sales
- Volume discounts available via committed contracts
Compare Anyscale vs Alternatives
Before committing to Anyscale, compare pricing with these 3 alternatives in the same category.
China-market apps and Chinese-first workloads
Full comparisonTesting Cerebras's unique speed advantage
Full comparisonLong-context (1M tokens) and Chinese-language apps
Full comparisonWhat Companies Actually Pay for Anyscale
How Anyscale Pricing Compares
| Software | Starting Price | Top Price |
|---|---|---|
| Anyscale | Free | $4.9591 per million tokens |
| Amazon Bedrock | $0.07 per million tokens | $75 per million tokens |
| Baidu ERNIE API | $0.1 per million tokens | $10 per million tokens |
| Cerebras Inference API | $0.1 per million tokens | $6 per million tokens |
| Claude API | $0.1 per million tokens | $75 per million tokens |
| Cloudflare Workers AI | Free | $4.881 per million tokens |
Detailed pricing comparisons:
How to Negotiate Anyscale Pricing
Anyscale contracts are negotiable. These 3 tactics are sourced from real buyer experiences and procurement specialists.
For committed contracts, buyers can negotiate for volume discounts by committing to a certain level of spending.
https://g2.comAnyscale's platform allows for efficient GPU utilization, auto-scaling, and the use of spot instances which can be 50-80% cheaper.
https://lawinsider.comBuyers should push for detailed breakdowns of all potential costs, including those beyond direct compute, as Anyscale's pricing structure can feel unclear.
https://g2.comAnyscale Price History
Pricing changes CostBench has tracked for Anyscale, verified across 2 snapshots going back to 2026Q2.
-
Anyscale tier renamed
"Anyscale Endpoints" renamed to "Pay As You Go"
-
Anyscale tier renamed
"Managed Ray Clusters" renamed to "Committed Contracts (BYOC)"
View Anyscale's full price history →
Anyscale Pricing FAQ
01 How much do Anyscale Endpoints cost?
Anyscale Endpoints charge per token. Small models like Llama 3.1 8B and Mistral 7B cost $0.15 per million tokens (input and output same rate). Llama 3.1 70B costs $1.00/M tokens. Llama 3.1 405B costs $5.00/M tokens. New accounts get $10 in free credits.
02 What is Anyscale built on?
Anyscale is built on Ray, an open-source distributed computing framework developed at UC Berkeley. Anyscale provides the managed, enterprise version of Ray with production-grade SLAs, managed clusters, and hosted inference endpoints.
03 Does Anyscale have a free tier?
Anyscale gives new accounts $10 in free credits for Endpoints usage. There is no permanently free tier — after credits are used, standard per-token rates apply.
04 Anyscale vs Together AI: which should I use?
Together AI and Anyscale both offer OpenAI-compatible inference for open-source models. Together AI has broader model selection and slightly more competitive pricing for commodity models. Anyscale is better if you're already using Ray for training or need the Ray ecosystem integration.
Is this pricing incorrect? — we'll verify and update it.