Skip to main content

Every frontier AI model, one subscription

Create a ClawAI account and reach Claude Opus 5, GPT-5, Gemini 3 Pro, Kimi K2, GLM-5.1, Qwen3, DeepSeek V3.2, Grok 4 and Amazon Bedrock from a single chat — no separate vendor accounts, no separate bills, no juggling API keys. Pick a plan, log in, and start working.

Start free and upgrade whenever you need more. Paid plans from $5 a month, cancel any time.

Last reviewed 2026-07-27

The models you get

One subscription covers every provider below. Switch models mid-conversation, or let ClawAI pick the best one for each message.

Newest models in the catalog

A concise snapshot of recently added models available through current connectors.

  • GLM-5.2Zhipu
  • Kimi K3Moonshot
  • Qwen3.5Alibaba
  • MiniMax M2.7MiniMax
  • DeepSeek V4 ProDeepSeek
  • DeepSeek V4 FlashDeepSeek
  • Gemini 3 Flash PreviewGoogle
  • Nemotron 3 SuperNVIDIA

Anthropic Claude

Deep reasoning, long documents, and careful code review.

  • Claude Fable 5
  • Claude Opus 5
  • Claude Sonnet 5
  • Claude Haiku 4.5

OpenAI GPT

Broad general capability with strong tool use and structured output.

  • GPT-5.6 Sol
  • GPT-5.6 Terra
  • GPT-5.6 Luna
  • GPT-5.5

Google Gemini

Very large context windows, fast responses, and native multimodal input.

  • Gemini 3.6 Flash
  • Gemini 3.5 Flash
  • Gemini 3.5 Flash-Lite
  • Gemini 3.1 Pro Preview

Moonshot Kimi

Long-context analysis and agentic workflows at low cost.

  • Kimi K3
  • Kimi K2.7 Code
  • Kimi K2.6

Zhipu GLM

Strong bilingual reasoning with explicit thinking modes.

  • GLM-5.2
  • GLM-5.1
  • GLM-5 Turbo
  • GLM-5

Alibaba Qwen

Excellent code generation and a wide range of model sizes.

  • Qwen3.7 Plus
  • Qwen3.6 Plus
  • Qwen3.5 Plus
  • Qwen3 Coder Next

DeepSeek

Mathematics, algorithms, and step-by-step reasoning.

  • DeepSeek V4 Pro
  • DeepSeek V4 Flash

xAI Grok

Fast conversational answers with current-events awareness.

  • Grok 4.5
  • Grok 4.20
  • Grok 4.3

Amazon Bedrock

Enterprise-grade hosting for teams already invested in AWS.

  • Nova 2 Lite
  • GPT-5.6 Sol
  • Claude Opus 5
  • Grok 4.3

New frontier models are added as they launch — your plan covers them from day one, metered against a single allowance.

Plans and pricing

Every paid plan unlocks every model. The only difference is how much you can use each day and each month.

Free

Try every frontier model with a small daily allowance.

$0.00

Daily tokens
10,000
Weekly token quota
20,000
Monthly tokens
Unlimited
Chats per day
5
Max messages per day
250
Max workspace connections
5
Max context packs
10
Max memory items
10
  • Compare mode
  • Judge mode
  • Research mode
  • Critic review
  • Workspaces
  • Memory
  • Context packs
  • Consensus Mode
  • Escalation Chain
  • Repair Lab
  • Task Decomposer
  • Best-of-N Generation
  • Verifier
  • Pipeline Lab
  • Cost-Aware Ensemble
  • Role Pack

Trial

A trial subscription, enable all features for trial 30 days

$2.00/month

Daily tokens
20,000
Weekly token quota
Unlimited
Monthly tokens
600,000
Chats per day
10
Max messages per day
100
Max workspace connections
10
Max context packs
25
Max memory items
25
  • Compare mode
  • Judge mode
  • Research mode
  • Critic review
  • Workspaces
  • Memory
  • Context packs
  • Consensus Mode
  • Escalation Chain
  • Repair Lab
  • Task Decomposer
  • Best-of-N Generation
  • Verifier
  • Pipeline Lab
  • Cost-Aware Ensemble
  • Role Pack

Starter

Everyday access to fast, low-cost frontier models.

$5.00/month

Daily tokens
50,000
Weekly token quota
250,000
Monthly tokens
750,000
Chats per day
10
Max messages per day
100
Max workspace connections
1
Max context packs
5
Max memory items
50
  • Compare mode
  • Judge mode
  • Research mode
  • Critic review
  • Workspaces
  • Memory
  • Context packs
  • Consensus Mode
  • Escalation Chain
  • Repair Lab
  • Task Decomposer
  • Best-of-N Generation
  • Verifier
  • Pipeline Lab
  • Cost-Aware Ensemble
  • Role Pack

Plus

More allowance and access to standard-tier models.

$8.00/month

Daily tokens
100,000
Weekly token quota
600,000
Monthly tokens
1,750,000
Chats per day
25
Max messages per day
250
Max workspace connections
2
Max context packs
15
Max memory items
200
  • Compare mode
  • Judge mode
  • Research mode
  • Critic review
  • Workspaces
  • Memory
  • Context packs
  • Consensus Mode
  • Escalation Chain
  • Repair Lab
  • Task Decomposer
  • Best-of-N Generation
  • Verifier
  • Pipeline Lab
  • Cost-Aware Ensemble
  • Role Pack
Most popular

Pro

Premium models, deep reasoning and heavy orchestration.

$15.00/month

Daily tokens
250,000
Weekly token quota
1,500,000
Monthly tokens
4,000,000
Chats per day
75
Max messages per day
750
Max workspace connections
5
Max context packs
50
Max memory items
1,000
  • Compare mode
  • Judge mode
  • Research mode
  • Critic review
  • Workspaces
  • Memory
  • Context packs
  • Consensus Mode
  • Escalation Chain
  • Repair Lab
  • Task Decomposer
  • Best-of-N Generation
  • Verifier
  • Pipeline Lab
  • Cost-Aware Ensemble
  • Role Pack

Team

Shared workspaces and a large pooled allowance.

$25.00/month

Daily tokens
750,000
Weekly token quota
4,000,000
Monthly tokens
11,000,000
Chats per day
250
Max messages per day
2,500
Max workspace connections
15
Max context packs
200
Max memory items
5,000
  • Compare mode
  • Judge mode
  • Research mode
  • Critic review
  • Workspaces
  • Memory
  • Context packs
  • Consensus Mode
  • Escalation Chain
  • Repair Lab
  • Task Decomposer
  • Best-of-N Generation
  • Verifier
  • Pipeline Lab
  • Cost-Aware Ensemble
  • Role Pack

Scale

High-volume access across every model class.

$50.00/month

Daily tokens
1,500,000
Weekly token quota
9,000,000
Monthly tokens
24,000,000
Chats per day
1,000
Max messages per day
10,000
Max workspace connections
50
Max context packs
1,000
Max memory items
25,000
  • Compare mode
  • Judge mode
  • Research mode
  • Critic review
  • Workspaces
  • Memory
  • Context packs
  • Consensus Mode
  • Escalation Chain
  • Repair Lab
  • Task Decomposer
  • Best-of-N Generation
  • Verifier
  • Pipeline Lab
  • Cost-Aware Ensemble
  • Role Pack

Unlimited

Unlimited chats and messages, with fair-use on premium cloud.

$200.00/month

Daily tokens
5,000,000
Weekly token quota
30,000,000
Monthly tokens
Unlimited
Chats per day
Unlimited
Max messages per day
Unlimited
Max workspace connections
200
Max context packs
5,000
Max memory items
100,000
  • Compare mode
  • Judge mode
  • Research mode
  • Critic review
  • Workspaces
  • Memory
  • Context packs
  • Consensus Mode
  • Escalation Chain
  • Repair Lab
  • Task Decomposer
  • Best-of-N Generation
  • Verifier
  • Pipeline Lab
  • Cost-Aware Ensemble
  • Role Pack

What happens when you send a message

  1. You send a message in a conversation, optionally attaching files or picking a specific model.

  2. Routing picks the model based on the mode you chose, what the message asks for, and which providers are healthy right now.

  3. Context is assembled saved memories, attached context packs and the relevant chunks of your uploaded files are merged into the prompt within a token budget, alongside the conversation history.

  4. The model answers streamed back as it is generated, with live timing and reasoning where the model exposes it, and an automatic fallback to another model if the first one fails.

  5. The exchange is recorded facts and preferences worth keeping are saved for next time, and the cost is metered against your allowance so there is no surprise at the end of the month.

Memory and context are switches you control per conversation. Everything a model was given is listed in a receipt attached to the answer, so you can check what informed it. Uploaded files are chunked and retrieved the same way, which is how ClawAI answers questions about your own documents instead of guessing from training data.

Beyond a single chat window

When one model’s answer isn’t enough on its own, ClawAI offers several ways to combine, check, and improve on it.

Parallel compare
send one prompt to two to five models at once and see every response side by side, with per-model latency and token counts.
Consensus
run the same prompt across several models and synthesize a single answer from where they agree.
Escalation chains
start with a fast, cheap model and automatically escalate to a stronger one when quality falls short of a threshold.
Best-of-N
generate several candidate answers and select the strongest one by a scoring pass.
Answer repair
detect and correct specific classes of errors in a prior response rather than regenerating from scratch.
Verification
have a second pass check a response for correctness before it reaches you, with configurable revision limits.
Role packs
run a prompt through a small team of role-specialized models (for example, a planner, a critic, and a writer) that hand off to one another.
Pipelines
chain multiple orchestration stages into a single named, repeatable workflow.
Judge & Critic review
an optional independent model reviews and scores a generated response before it is treated as final.

In your editor

ClawAI works inside VS Code

Every model on your subscription, in the editor you already use. The extension is a thin client — your account, quotas and history stay on the platform, so a conversation started in the browser continues in the editor.

No second subscription
It uses the ClawAI account you already have and draws on the same allowance. Nothing extra to buy.
No API keys in your editor
Routing happens on the platform, so the extension never holds a model provider's credential.
Works with self-hosting
Point it at ClawAI's hosted platform or at your own deployment. You choose at sign-in.

For organisations

Need ClawAI inside your own network?

Companies can have ClawAI deployed on their own servers running local models only, so no prompt, document or conversation ever leaves their infrastructure. This is a bespoke deployment we set up with you — talk to us and we will scope it.

Runs on your servers
Deployed into your data centre or private cloud, managed by your team, behind your firewall.
Local models only
Open-weight models served on your own hardware. No external provider calls and no third-party data processing.
Your controls, your audit trail
SSO, role-based access, retention rules and a full audit log of every request, all kept inside your environment.

Start on the free plan

Create an account and send your first message in under a minute. Upgrade only when you outgrow the free allowance.