Developer Pricing
Build with Riven. Simple monthly plans, predictable costs.
Every plan includes a monthly token allocation shared across input and output — no per-call billing, no surprise invoices. Pick a plan, get your API key, and start building.
Looking for Chat, Relay, or team plans instead? See consumer pricing. Prefer pay-as-you-go? See platform.rivenai.io/pricing.
Plans
Every plan includes a monthly token allocation.
Pay-As-You-Go
- Pay only for what you use
- $3.00 per 1M tokens
- All Riven models
- No monthly minimum
- OpenAI-compatible API
Riven Pro
- Single seat
- All standard models
- Priority queue
- Memory
- OpenAI-compatible API
- 3M tokens/month included
- $5/mo bundled API credit
- 200 agentic steps/mo
- CLI + platform console access
Riven Team
- Multi-seat
- Per-seat billing
- All standard models
- Shared memory
- Admin console
- 4M tokens/month/seat included
- $10/seat/mo bundled API credit
- 1000 agentic steps/seat/mo
- Team admin + audit
How it works
How allocation works
No per-token invoices. Each plan grants a fixed pool of tokens (input + output combined) every month.
Free
100 requests per day across all Riven models — enough to explore and prototype.
Riven Pro
100,000 tokens per month plus a $5 bundled API credit for individual builders.
Pay-As-You-Go
$3.00 per 1M tokens, no monthly minimum — scale past any plan cap without upgrading.
Models
Every plan includes access to the full model catalog.
| Model | Provider | Context window | Capabilities |
|---|---|---|---|
| claude-fable-5 | anthropic | — | chat, reasoning |
| claude-opus-4-7 | anthropic | 200,000 | chat, reasoning |
| claude-opus-4-8 | anthropic | 200,000 | chat, reasoning |
| claude-sonnet-4-6 | anthropic | 200,000 | chat, reasoning |
| claude-haiku-4-5 | anthropic | 200,000 | chat |
| gpt-5 | openai | 128,000 | chat, reasoning |
| gpt-5.4-mini | openai | 128,000 | chat |
| gpt-5.4-nano | openai | 128,000 | chat |
| gpt-5.5 | openai | 256,000 | chat, reasoning |
| gpt-5.6 | openai | — | chat, reasoning |
| gpt-5.6-luna | openai | — | chat, reasoning |
| gpt-5.6-sol | openai | — | chat, reasoning |
| gpt-5.6-terra | openai | — | chat, reasoning |
| gpt-4o | openai | — | chat, reasoning |
| gpt-4o-mini | openai | — | chat |
| kimi-k2.5 | moonshot | 256,000 | chat |
| kimi-k2.6 | moonshot | — | chat |
| kimi-k2.7-code | moonshot | — | chat, code, reasoning |
| kimi-k3 | moonshot | — | chat |
| codestral-2508 | mistral | — | chat, code |
| mistral-large-2512 | mistral | — | chat |
| mistral-large-latest | mistral | 256,000 | chat |
| mistral-medium-3 | mistral | — | chat |
| mistral-medium-3.5 | mistral | — | chat, reasoning |
| mistral-medium-latest | mistral | 128,000 | chat |
| mistral-small-latest | mistral | 128,000 | chat |
| sonar | perplexity | 128,000 | chat |
| sonar-pro | perplexity | 128,000 | chat |
| sonar-reasoning-pro | perplexity | 128,000 | chat, reasoning |
| llama-3.1-8b-instant | groq | — | chat |
| llama-3.3-70b-versatile | groq | — | chat |
| zai-glm-4.7 | cerebras | — | chat |
| kimi-k2.7-code-highspeed | moonshot | — | chat, code, reasoning |
| gpt-oss-120b | fireworks | — | chat |
| glm-5.2 | Ultra-native | — | chat, reasoning |
| qwen3.6-35b-a3b-fp8 | Ultra-native | 131,072 | chat, reasoning |
| nomic-embed-text | Ultra-native | 2,048 | embeddings |
| rvn-assistant-v2 | Ultra-native | 128,000 | chat |
| riven-instant | Ultra-native | — | chat |
| riven-fast | Ultra-native | — | chat |
| riven-core | Ultra-native | — | chat |
| riven-pro | Ultra-native | — | chat, reasoning |
| riven-max | Ultra-native | — | chat, reasoning |
| riven-research | Ultra-native | — | chat, reasoning |
| riven-reason | Ultra-native | — | chat, reasoning |
| riven-code | Ultra-native | — | chat, code |
| riven-agent | Ultra-native | — | chat |
| riven-vision | Ultra-native | — | chat, vision |
| riven-image | Ultra-native | — | |
| riven-embed | Ultra-native | — | embeddings |
| riven-embed-fast | Ultra-native | — | embeddings |
| riven-chat-sm | Ultra-native | — | chat |
| riven-chat-md | Ultra-native | — | chat |
| riven-chat-lg | Ultra-native | — | chat |
| riven-chat-xl | Ultra-native | — | chat |
What if I need more?
For pay-as-you-go pricing, higher limits, SSO, or on-premise deployment, use the Platform console — a separate PAYG + Enterprise surface.
- Pay-as-you-go token pricing
- SAML / Entra ID SSO
- RBAC and audit logs
- Private model deploys
- Dedicated support + SLA
- On-premise deployment
Questions
How does the monthly allocation work?
Each plan includes a fixed number of tokens (input + output combined) per calendar month. Your allocation resets automatically on the 1st of every month, UTC.
What happens if I use my full allocation?
API calls are paused until your next monthly reset, or you can upgrade to a higher tier instantly for more headroom.
Do unused tokens roll over?
No. Allocations reset to the plan amount every month — they do not accumulate or carry over.
Can I change plans anytime?
Yes. Upgrade or downgrade at any time from the console Plan page. Upgrades apply immediately.
Prefer pay-as-you-go?
For token-metered PAYG billing with no monthly commitment, use platform.rivenai.io instead of a dev subscription.