Agents that get work done.
One bill. No surprises.

Pick a plan, sign up, and start building. Tokens are included by default. Advanced users can bring their own LLM provider or train custom models.

Pick a plan. Usage included.

Every plan ships the whole platform - runtime, hub, dev toolkit - with a generous monthly usage allowance baked in. Plans are sized against each other, so you pick a size rather than doing arithmetic. Advanced users can bring their own LLM provider and pay them directly.

All prices VAT-inclusive (gross). EU B2B with valid VAT ID and non-EU customers see the net price.

Save ~5% with yearly billing
Free
No card required
$0
 

Try it out. See if it fits your work.

  • up to 5 agents
  • Light usage - to try things out
  • 1 GB workspace / mo
  • 1 GB hub storage / mo
  • 30 min web browsing / mo
  • Community support
Standard
Save ~5% billed yearly
$19/month
$19/mo, or $18/mo billed annually.

For personal projects and occasional automation.

  • up to 100 agents
  • Everyday usage
  • 5 GB workspace / mo
  • 10 GB hub storage / mo
  • 5 hours web browsing / mo
  • 180 compute credits (3 GPU-hours) / mo
  • Agent user self-registration
  • Email support
Most popular
Pro
Save ~6% billed yearly
$96/month
$96/mo, or $90/mo billed annually.

For daily workflows and serious automation.

  • up to 500 agents
  • 8x Standard's usage
  • 20 GB workspace / mo
  • 50 GB hub storage / mo
  • 15 hours web browsing / mo
  • 900 compute credits (15 GPU-hours) / mo
  • Agent user self-registration
  • Agent monetization
  • Priority email support
Max
Save ~5% billed yearly
$190/month
$190/mo, or $180/mo billed annually.

For power users, small businesses, and multi-agent setups.

  • up to 1,000 agents
  • 20x Standard's usage
  • 50 GB workspace / mo
  • 200 GB hub storage / mo
  • 40 hours web browsing / mo
  • 2,400 compute credits (40 GPU-hours) / mo
  • Agent user self-registration
  • Agent monetization
  • Priority support

Two frontier model levels on every plan: Surogate (Sonnet class), and Surogate Pro (Opus class). Pro is stronger, and draws ~2.5x faster from the plan allowance.

Enterprise

Let's talk.

Custom usage volume. Dedicated compute. SSO, audit logs, RBAC. SLA, dedicated support, custom contracts. For teams of 5+, regulated industries, and ML teams shipping production models.

Contact sales

Need more usage or browser time? You can always Top Up your wallet

Bring your own model. Train your own model.

Available on every paid plan. Pick what you need, configure in your account settings. No GPU markup from us, no double-dipping.

01

Use your own LLM

Connect any OpenAI-compatible provider. You pay them directly. Cancel the bundled usage portion of your plan in settings to save the cost.

External providers
OpenRouterOpenAIAnthropicTogetherGroq
Self-hosted
vLLMTGIOllamaAny OpenAI-compatible API
Your fine-tuned models
Deploy via dstack or skypilot
02

Fine-tune your own models

Train on your own data using dstack or skypilot. Your cloud, your GPUs, your bill - we orchestrate the run, capture logs and metrics, and track lineage in the hub.

Cloud GPU providers
AWSGCPLambdaRunPodAnywhere
03

Datasets & evaluation

Generate synthetic datasets at scale. Run standardized or custom eval suites. Compute credits work like the other wallet resources - plan grant resets monthly, top-ups roll over. One credit is one minute on the baseline GPU tier, so faster GPUs simply draw proportionally more.

HumanEvalMBPPterminal-benchSWE-benchCustom suites
Standard180 credits - 3 GPU-hours
Pro900 credits - 15 GPU-hours
Max2,400 credits - 40 GPU-hours
Top-up$2.00 / GPU-hour - rolls over
04

Internal hub

A git-backed registry for your models, datasets, and checkpoints - like a private Hugging Face Hub. Public read for community sharing, private read/write for your work. Backed by Cloudflare R2 - no egress fees.

Free1 GB hub storage
Standard10 GB hub storage
Pro50 GB hub storage
Max200 GB hub storage
Extra+$0.0605 / GB / month - recurring
07FAQ

Questions, answered in plain language.

Can't find what you're looking for? Talk to a human →

Can I switch plans anytime?

Yes. Upgrades take effect immediately and we prorate the difference. Downgrades take effect at the next billing cycle. No fees either way.

Is usage really included? What about overages?

Every paid plan ships with a monthly usage allowance baked into the price - Standard is the baseline, Pro gives you 8x that, Max 20x - all served on frontier-grade models. There are no surprise overage bills: when your monthly allowance runs out, agents pause until you top up your wallet (rolling-over cash, from $2) or wait for the next cycle. You stay in control.

What happens if my usage runs out mid-task?

The operation stops. No auto-upgrade, no overdraft, no end-of-month invoice. Top up the wallet with as little or as much as you want - top-up balances never expire until you use them - and the task resumes. You can also bring your own LLM provider and pay them directly, which bypasses our usage allowance entirely.

How do top-ups work?

Each consumable resource - usage, web browsing time, compute credits - works like a prepaid wallet. Your plan grants a monthly balance (resets each cycle, lose-it-or-use-it). Top-ups add more on demand, and that cash rolls over as long as your account is active. When the wallet hits zero, the operation stops.

Can I bring my own LLM and skip paying for usage?

Yes. Connect any OpenAI-compatible provider - OpenRouter, OpenAI, Anthropic, Together, Groq, vLLM, TGI, Ollama, or your own fine-tuned deployment. You pay them directly. In settings, cancel the bundled usage portion of your plan to save the cost.

Can I share my plan with my team?

The Standard, Pro, and Max plans are single-user. For team use, the Enterprise plan supports multiple seats, shared resources, role-based access, SSO, and audit logs. Custom team pricing is available - talk to sales.

What about VAT?

All prices on this page are VAT-inclusive (gross). EU B2B customers with a valid VAT ID and non-EU customers see the net price - we don’t charge VAT in those cases.

What about data privacy?

Your workspace is private and yours. Your hub artifacts are private by default. We don’t train on your data. You can export everything anytime, with no egress fees on hub content.

Do you charge for failed agent runs?

Tokens consumed by the model are billed whether the agent succeeded or not - same as every other LLM-based service. We don’t charge browser time or extra compute for failures on our end.

Is there a free trial of paid plans?

The Free plan is open-ended - use it as long as you like, with a starter usage allowance or BYO LLM from day one. If a paid plan doesn’t work for you in the first 14 days, we’ll refund it on request, no questions asked.

Can I run an agent 24/7?

Yes. Agents run continuously within your plan’s concurrent-agent limit - usage and web browsing hours are the wallets to watch for long-running tasks. The dashboard shows live consumption so you can tune behavior or top up as needed.

How does fine-tuning work? Do I pay for GPUs?

Train on your own data out of the box on our managed GPU cloud - pick a GPU, hit start, and the run draws from your compute credits. No cloud account needed. Prefer your own infrastructure? Connect your cloud (AWS, GCP, Lambda, RunPod, or your own Modal account) via dstack or skypilot and train there instead: your GPUs, your bill, zero credit usage and no GPU markup from us. Either way we orchestrate the run, capture logs and metrics, and track lineage in the hub.

How are compute credits different from agent usage?

Compute credits meter GPU time - fine-tuning, synthetic dataset generation, and eval runs (HumanEval, MBPP, terminal-bench, SWE-bench, custom suites) - while your usage allowance covers the LLM calls your agents make. One credit is one minute on the baseline GPU tier, so faster GPUs draw proportionally more (an H100 minute is about 6.7 credits) and you pay for exactly the hardware you picked. Standard includes 180 credits/month (3 GPU-hours), Pro 900 (15 GPU-hours), Max 2,400 (40 GPU-hours). Top-up rate is $2.00 per GPU-hour and rolls over. Training on your own connected cloud never touches the wallet.

Why bring my own cloud instead of using yours?

You probably already have credits, a preferred region, specific compliance needs, or a GPU provider relationship. Our managed GPU cloud is the default so you can start without a cloud account, but we won’t lock you into it - connect your own and your provider bills you directly, with no GPU markup from us and no credits consumed. We’d rather make money on the platform than on reselling compute.

Can I serve my fine-tuned models?

Yes. Deploy via your own cloud (dstack/skypilot), and we orchestrate the endpoint. Your agents can call your custom models directly - no usage charges from us when you use them.

What’s the hub for?

A git-backed registry for your models, datasets, and checkpoints - like a private Hugging Face Hub. Public read for community sharing, private read/write for your work. Backed by Cloudflare R2 with no egress fees. Private storage scales with your plan: 1 GB on Free, 10 GB on Standard, 50 GB on Pro, 200 GB on Max. Need more on any plan? Extra hub storage is a recurring add-on at +$0.0605/GB/month.

What if I need more than Max but Enterprise is overkill?

Stack add-ons. A Max plan with extra workspace storage (+$0.0605/GB/mo), extra hub storage (+$0.0605/GB/mo), and ongoing wallet top-ups covers a lot of ground before you hit Enterprise territory. Talk to us if you’re not sure.

08Ready to start?

Build the agent you actually want.

Start free, no credit card. Pick a plan when you need more agents, more usage, or your own fine-tuned models. Bring your own LLM whenever you like.