Simple Pricing

One API.
100+ LLMs.
Smart Routing.

Start free. Scale on Grizzly Pro $99 — built for agents. Services optional.

Pay-as-you-go

Add Credits

$10

minimum

Add credits as you need them. No subscription - pay only for what you use.

  • $10 minimum top-up
  • Add credits anytime, $10–$10,000
  • Credits never expire
  • Intelligent routing across 104+ models
  • OpenAI, Anthropic, Google, Meta, DeepSeek & more
  • Community support
  • Up to 5 API keys
Best for agents

Grizzly Pro

$99

/month

Best for OpenClaw, Hermes, and multi-agent stacks. Caching, guardrails, budgets, 100 keys.

  • 25M tokens included monthly
  • Semantic caching (up to 30% cost savings)
  • Guardrails & safety filtering
  • Advanced analytics dashboard
  • Cost optimizer recommendations
  • Virtual keys with budgets
  • Up to 100 API keys
  • Priority email support
Optional

Services

Custom

Need engineers embedded? SMB builds or enterprise FDE — only if the API alone isn’t enough.

  • Everything in Grizzly Pro, plus:
  • Dedicated engineering lead + named team
  • Custom RAG pipelines over your data
  • Private model fine-tuning & evaluation
  • SSO, RBAC, audit logs, SOC2-aligned deploy
  • Custom guardrails & policy enforcement
  • Production on-call + 99.9% SLA
  • Quarterly business reviews

Optional services

Just need the API? You’re done above.

Want embedded engineers for a custom build? SMB fixed-scope or enterprise FDE — separate from Pro usage.

Token Pricing

We add a small markup over provider costs. See live pricing across all models on our Models page.

Input Tokens

$0.04

per 1M tokens

vs OpenAI $2.50/1M

Output Tokens

$0.12

per 1M tokens

vs OpenAI $10/1M

Intelligent Routing

Every call

to the most efficient model

simple tasks → Llama 3B · reasoning → Claude

Feature Comparison

Compare every tier side-by-side. All plans include intelligent routing, semantic caching, and our unified API.

FeatureAdd CreditsGrizzly ProServices
Monthly tokensPay-as-you-go25MProject-scoped
API keys5100Unlimited
Intelligent routing
Semantic caching
Guardrails & safety
Analytics dashboardAdvancedCustom
Virtual keys & budgets
Custom RAG pipelines over your data
Private fine-tuning & eval suite
Custom guardrails & policy
SSO, RBAC, audit logs
SOC2-aligned deployment
Budget alerts
Dedicated engineering lead
Working software shipped
Production hardening + handoff
99.9% SLA + on-call rotation
Quarterly business reviews
SupportCommunityPriority emailDedicated FDE team

Frequently Asked Questions

How does intelligent routing work?

Our routing layer analyzes each request and picks the most efficient model for the job. Simple tasks like summarization go to fast, low-cost models (starting at $0.04/1M), while complex reasoning goes to high-reasoning models like Claude 70B or Grizzly 1.0. You always get the best quality-to-cost ratio.

What happens to my credits?

Credits never expire. Grizzly Pro includes a monthly token allocation. Optional services (SMB/enterprise builds) are quoted separately after discovery.

How does Grizzly Pro compare for agents?

Grizzly Pro is $99/month with 25M routed tokens across the catalog — OpenAI, Anthropic, Google, Meta, xAI, DeepSeek, and more. Semantic caching, guardrails, analytics, virtual keys with budgets, and up to 100 API keys. Built for multi-agent stacks that burn tokens all day.

Do I need services / FDE?

Most teams only need the API + Pro. Choose SMB or Enterprise services when you want engineers embedded for a custom build, RAG, or private deploy. Details on /for-smb and /enterprise.

Can I switch plans?

Yes, you can upgrade or downgrade anytime. Upgrades take effect immediately. Downgrades take effect at the start of your next billing cycle.

What models are available?

Access to 104+ models across OpenAI, Anthropic, Google, Meta, DeepSeek, Mistral, Groq, Fireworks, Together.ai, and more. Full pricing list on our Models page.

Route every LLM call through one API

Get a free API key. No credit card required. Start routing every LLM call automatically.