Try it out. See if it fits your work.
- up to 5 agents
- Light usage - to try things out
- 1 GB workspace / mo
- 1 GB hub storage / mo
- 30 min web browsing / mo
- Community support
Pick a plan, sign up, and start building. Tokens are included by default. Advanced users can bring their own LLM provider or train custom models.
Every plan ships the whole platform - runtime, hub, dev toolkit - with a generous monthly usage allowance baked in. Plans are sized against each other, so you pick a size rather than doing arithmetic. Advanced users can bring their own LLM provider and pay them directly.
All prices VAT-inclusive (gross). EU B2B with valid VAT ID and non-EU customers see the net price.
Try it out. See if it fits your work.
For personal projects and occasional automation.
For daily workflows and serious automation.
For power users, small businesses, and multi-agent setups.
Two frontier model levels on every plan: Surogate (Sonnet class), and Surogate Pro (Opus class). Pro is stronger, and draws ~2.5x faster from the plan allowance.
Custom usage volume. Dedicated compute. SSO, audit logs, RBAC. SLA, dedicated support, custom contracts. For teams of 5+, regulated industries, and ML teams shipping production models.
Contact sales →Need more usage or browser time? You can always Top Up your wallet
Available on every paid plan. Pick what you need, configure in your account settings. No GPU markup from us, no double-dipping.
Connect any OpenAI-compatible provider. You pay them directly. Cancel the bundled usage portion of your plan in settings to save the cost.
Train on your own data using dstack or skypilot. Your cloud, your GPUs, your bill - we orchestrate the run, capture logs and metrics, and track lineage in the hub.
Generate synthetic datasets at scale. Run standardized or custom eval suites. Compute credits work like the other wallet resources - plan grant resets monthly, top-ups roll over. One credit is one minute on the baseline GPU tier, so faster GPUs simply draw proportionally more.
A git-backed registry for your models, datasets, and checkpoints - like a private Hugging Face Hub. Public read for community sharing, private read/write for your work. Backed by Cloudflare R2 - no egress fees.
Yes. Upgrades take effect immediately and we prorate the difference. Downgrades take effect at the next billing cycle. No fees either way.
Every paid plan ships with a monthly usage allowance baked into the price - Standard is the baseline, Pro gives you 8x that, Max 20x - all served on frontier-grade models. There are no surprise overage bills: when your monthly allowance runs out, agents pause until you top up your wallet (rolling-over cash, from $2) or wait for the next cycle. You stay in control.
The operation stops. No auto-upgrade, no overdraft, no end-of-month invoice. Top up the wallet with as little or as much as you want - top-up balances never expire until you use them - and the task resumes. You can also bring your own LLM provider and pay them directly, which bypasses our usage allowance entirely.
Each consumable resource - usage, web browsing time, compute credits - works like a prepaid wallet. Your plan grants a monthly balance (resets each cycle, lose-it-or-use-it). Top-ups add more on demand, and that cash rolls over as long as your account is active. When the wallet hits zero, the operation stops.
Yes. Connect any OpenAI-compatible provider - OpenRouter, OpenAI, Anthropic, Together, Groq, vLLM, TGI, Ollama, or your own fine-tuned deployment. You pay them directly. In settings, cancel the bundled usage portion of your plan to save the cost.
The Standard, Pro, and Max plans are single-user. For team use, the Enterprise plan supports multiple seats, shared resources, role-based access, SSO, and audit logs. Custom team pricing is available - talk to sales.
All prices on this page are VAT-inclusive (gross). EU B2B customers with a valid VAT ID and non-EU customers see the net price - we don’t charge VAT in those cases.
Your workspace is private and yours. Your hub artifacts are private by default. We don’t train on your data. You can export everything anytime, with no egress fees on hub content.
Tokens consumed by the model are billed whether the agent succeeded or not - same as every other LLM-based service. We don’t charge browser time or extra compute for failures on our end.
The Free plan is open-ended - use it as long as you like, with a starter usage allowance or BYO LLM from day one. If a paid plan doesn’t work for you in the first 14 days, we’ll refund it on request, no questions asked.
Yes. Agents run continuously within your plan’s concurrent-agent limit - usage and web browsing hours are the wallets to watch for long-running tasks. The dashboard shows live consumption so you can tune behavior or top up as needed.
Train on your own data out of the box on our managed GPU cloud - pick a GPU, hit start, and the run draws from your compute credits. No cloud account needed. Prefer your own infrastructure? Connect your cloud (AWS, GCP, Lambda, RunPod, or your own Modal account) via dstack or skypilot and train there instead: your GPUs, your bill, zero credit usage and no GPU markup from us. Either way we orchestrate the run, capture logs and metrics, and track lineage in the hub.
Compute credits meter GPU time - fine-tuning, synthetic dataset generation, and eval runs (HumanEval, MBPP, terminal-bench, SWE-bench, custom suites) - while your usage allowance covers the LLM calls your agents make. One credit is one minute on the baseline GPU tier, so faster GPUs draw proportionally more (an H100 minute is about 6.7 credits) and you pay for exactly the hardware you picked. Standard includes 180 credits/month (3 GPU-hours), Pro 900 (15 GPU-hours), Max 2,400 (40 GPU-hours). Top-up rate is $2.00 per GPU-hour and rolls over. Training on your own connected cloud never touches the wallet.
You probably already have credits, a preferred region, specific compliance needs, or a GPU provider relationship. Our managed GPU cloud is the default so you can start without a cloud account, but we won’t lock you into it - connect your own and your provider bills you directly, with no GPU markup from us and no credits consumed. We’d rather make money on the platform than on reselling compute.
Yes. Deploy via your own cloud (dstack/skypilot), and we orchestrate the endpoint. Your agents can call your custom models directly - no usage charges from us when you use them.
A git-backed registry for your models, datasets, and checkpoints - like a private Hugging Face Hub. Public read for community sharing, private read/write for your work. Backed by Cloudflare R2 with no egress fees. Private storage scales with your plan: 1 GB on Free, 10 GB on Standard, 50 GB on Pro, 200 GB on Max. Need more on any plan? Extra hub storage is a recurring add-on at +$0.0605/GB/month.
Stack add-ons. A Max plan with extra workspace storage (+$0.0605/GB/mo), extra hub storage (+$0.0605/GB/mo), and ongoing wallet top-ups covers a lot of ground before you hit Enterprise territory. Talk to us if you’re not sure.
Start free, no credit card. Pick a plan when you need more agents, more usage, or your own fine-tuned models. Bring your own LLM whenever you like.