Everything here takes minutes, not hours. No card, no email, no account to create on this website β download the Sandbox app for Windows, install it, and you're comparing models in under a minute. (macOS build is coming soon β join the waitlist on the pricing page.)
1 Β· Download and install (30 seconds)
- Click Download for Windows on any page. The installer is a small .exe β no admin rights required on most machines.
- Run it. No email, no password, no account to set up here β Sandbox opens straight to your workspace.
- Your trial starts automatically: 300 trial credits and full Pro-level access for 3 days, completely free.
2 Β· Turn on your models
Open Settings β Models. Every model in the directory is a toggle:
- Chat & text: GPT-5, o3, Claude Opus/Sonnet, Gemini 2.5 Pro/Flash, Grok 4, DeepSeek V3, Llama 4, Mistral Large.
- Image: Midjourney v7, DALLΒ·E 3, Imagen 4, Stable Diffusion 3.
- Voice & music: ElevenLabs voices, Suno v4.
- Video: Runway Gen-4.
Enabled models appear as lanes in your workspace. Starter plans include the 8 core models; Pro and above unlock the full directory.
3 Β· Run your first parallel comparison
- Type any real task into the prompt bar β the one you'd normally ask one chatbot.
- Press β/Ctrl + Enter to send it to all enabled models at once (or pick specific ones).
- Answers render side by side with latency, credit cost and a quality score per lane.
- Click Best answer to adopt one, or Merge to blend fragments from several into a final draft.
4 Β· Delegate to agents (Pro and up)
Open Agents and describe an outcome instead of a prompt:
- Code Agent β point it at a GitHub repo, describe the change, approve the plan, review the diff.
- Research Agent β brief a market or competitor study; receive a sourced report in the background.
- Media Agent β brief a video campaign; approve scripts and frames at the gate; receive rendered video.
- Copy Agent β brief a launch; receive multi-angle, multi-language copy scored to your brand voice.
Every agent run shows a plan before it starts, a spend cap you set, and a receipt after it finishes.
5 Β· Manage credits honestly
- Costs are per answer, not per subscription: a standard text answer is 1β3 credits; images 4; video 12 per 5 seconds. The exact price shows before anything expensive runs.
- Soft cap: set a monthly ceiling in Settings β heavy jobs pause at 90% instead of overspending.
- Rollover: Pro credits roll over one month; Ultra three months.
- Top-ups: $10 = 400 credits anytime, any plan, no plan change.
6 Β· Connect the API (Ultra and up)
Generate an API key under Settings β API. One endpoint, model chosen by parameter β drop-in compatible with OpenAI's format, so existing code changes to a base URL and a model name:
POST https://api.sandbox-ai.tech/v1/chat with "model": "claude-opus" β or "auto" to let routing pick the cheapest model that passes your quality bar. Webhooks deliver agent completions to your systems.
Trial timeline β what happens when
| Day | What happens |
|---|---|
| Day 0 | Full Pro access + 300 credits. No card, no email, no signup. |
| Day 3 | Trial ends. A payment link appears inside the app β pick a plan there, or ignore it and let it lapse. No dark patterns, no emails sent. |
| After Day 3 (no action) | Workspace pauses; nothing is charged; your history waits 90 days. |
| Any time later | Open the app, use the payment link, and everything resumes exactly where it stopped. |
Troubleshooting
- "Model busy" β shared lanes saturate at peak (rare, Pro+ has priority GPU). Retry, or route the prompt to the same model's Flash/Small variant.
- Slow video renders β Runway queues are longest Monday mornings EU time; storyboard first so renders happen once.
- Agent stopped at approval gate β that's by design: expensive steps wait for you. Check the gate in the Agents tab.
- Imported custom GPT behaves differently β presets run on the model you choose; some instructions behave differently across models. The comparison mode shows you exactly where.
Still reading? You could be comparing models already.
Install takes 30 seconds. The first comparison takes 10.
Next steps
- ChatGPT vs Claude vs Gemini β read before your first comparison
- Delegate your first real task to the Code Agent
- Calculate what consolidation saves you
More guides
- Install & uninstall β clean setup, updates, and removing Sandbox without losing your account
- Billing & plans β credits, upgrades, downgrades, invoices and VAT
- Invite your team β adding seats and permissions on Ultra and Custom
- Troubleshooting β fixes for the most common issues