Start free, scale as you grow. No vendor lock-in, no per-token surprises for standard calls.
FREE
Kick the tires. No card required.
STARTER
For developers shipping real projects.
PRO
For production apps and growing teams.
TEAM
For engineering teams with shared billing + SLA.
ENTERPRISE
Unlimited scale, dedicated infra, negotiated SLAs.
Full Comparison
Every capability, side by side.
| Feature | Free | Starter | Pro | Team | Enterprise |
|---|---|---|---|---|---|
| Usage & Quotas | |||||
| API requests / month | 100 | 1,000 | 8,000 | 40,000 | Custom |
| Tokens / month | 200K | 1.5M | 12M | 60M | Custom |
| Rate limit (req / min) | 10 | 60 | 200 | 600 | Custom |
| Concurrent requests | 2 | 10 | 30 | 100 | Custom |
| Models & Access | |||||
| Economy models | ✓ | ✓ | ✓ | ✓ | ✓ |
| Standard models | — | ✓ | ✓ | ✓ | ✓ |
| Premium models | — | Capped | ✓ | ✓ | ✓ |
| Flagship models cap / month | — | 10 | 60 | 250 | Custom |
| GPU Tunnel (local / self-hosted) | — | — | ✓ | ✓ | ✓ |
| Private model hosting | — | — | — | — | ✓ |
| 60+ model catalogue | Economy | Full | Full | Full | Full + Private |
| API & Connectivity | |||||
| API keys | 1 | 3 | 10 | 25 | Unlimited |
| OpenAI-compatible endpoints | ✓ | ✓ | ✓ | ✓ | ✓ |
| Streaming (SSE) | ✓ | ✓ | ✓ | ✓ | ✓ |
| Webhooks | — | — | ✓ | ✓ | ✓ |
| Custom base URL / BYOK | — | — | — | — | ✓ |
| SDK support (Python, Node, Go) | ✓ | ✓ | ✓ | ✓ | ✓ |
| RAG & Memory | |||||
| RAG indexing ops / month | — | 150 | 1,500 | 6,000 | Unlimited |
| RAG search ops / month | — | 800 | 8,000 | 35,000 | Unlimited |
| Max document size | — | 5 MB | 25 MB | 100 MB | Custom |
| Vector search (built-in) | — | ✓ | ✓ | ✓ | ✓ |
| Context-aware memory | — | — | ✓ | ✓ | ✓ |
| Analytics & Observability | |||||
| Usage dashboard | Basic | Standard | Full | Full | Custom |
| Per-model cost breakdown | — | — | ✓ | ✓ | ✓ |
| Latency & error traces | — | — | ✓ | ✓ | ✓ |
| Data export (CSV / JSON) | — | — | ✓ | ✓ | ✓ |
| Audit logs | — | — | — | ✓ | ✓ |
| Exportable audit trail | — | — | — | ✓ | ✓ |
| Security & Compliance | |||||
| TLS encryption in transit | ✓ | ✓ | ✓ | ✓ | ✓ |
| Dual-gate spend caps | ✓ | ✓ | ✓ | ✓ | ✓ |
| Role-based access (5-role RBAC) | — | — | — | ✓ | ✓ |
| SSO / SAML 2.0 | — | — | — | — | ✓ |
| VPC / private deployment | — | — | — | — | ✓ |
| Data retention controls | — | — | — | ✓ | ✓ |
| Encryption at rest | ✓ | ✓ | ✓ | ✓ | ✓ |
| Teams & Collaboration | |||||
| Seats | 1 | 1 | 1 | 5 (+$25/ea) | Unlimited |
| Shared request pool | — | — | — | ✓ | ✓ |
| Per-seat token budgets | — | — | — | ✓ | ✓ |
| Workspace isolation | — | — | — | ✓ | ✓ |
| White-label / custom branding | — | — | — | — | ✓ |
| Support & SLA | |||||
| Support channel | Community | Priority | Dedicated Slack | ||
| First response time | — | 24 h | 24 h | 4 h | 1 h |
| Uptime SLA | — | — | — | 99.9% | 99.95% |
| Named customer success manager | — | — | — | — | ✓ |
| White-glove onboarding | — | — | — | — | ✓ |
FAQ
Yes — plan changes take effect immediately. Downgrades apply at the end of your billing period. You keep your data either way.
Annual plans are billed once per year and include a ~20% discount compared to paying monthly. You can switch to annual at any time.
Each request to /api/v1/chat or /api/v1/ai counts as one API call. RAG chat at /api/v1/rag/chat counts as one API call plus one RAG search.
Requests are hard-blocked until the start of your next billing cycle. Upgrade to a higher plan at any time for more monthly capacity — there are no overage charges.
We offer a 14-day money-back guarantee on your first paid month. After that, refunds are handled case-by-case — reach out to support.
Yes — via the open-source self-hosted version on GitHub. The managed cloud product uses TARQA AI's infrastructure and keys are included in the plan price.
Free forever. No credit card required to start.