Features Pricing Docs About
PRICING

Simple pricing.
No token markups.

Self-host for free forever. Upgrade to Pro for the intelligent layer — auto-routing, semantic memory, and reasoning. Cloud API costs pass through to you at exact provider rates. No Dream-Weaver markup, ever.

Monthly
Annual
OPEN CORE
$0
free forever · self-hosted
Full OpenAI-compatible gateway. Local GPU inference. Community support. No credit card, no time limit.
  • ✓OpenAI-compatible API
  • ✓Unlimited local Ollama inference
  • ✓Streaming SSE + function calling
  • ✓Explicit model routing
  • ✓JSON mode + vision
  • ✓REST API
  • ✓Community support (GitHub Issues)
  • –Auto-routing (MoE classifier)
  • –Semantic memory (pgvector)
  • –LSR reasoning pod
  • –MCP tool server
  • –Multi-tenant API keys
ENTERPRISE
Custom
volume · annual contract
Air-gapped VPC deploy, SSO/SAML, SLA guarantee, dedicated support engineering, and custom model integrations.
  • ✓Everything in Pro
  • ✓Air-gapped / VPC deployment
  • ✓SSO / SAML / LDAP
  • ✓Dedicated routing cluster
  • ✓Custom model integrations
  • ✓SLA + 99.9% uptime guarantee
  • ✓White-glove onboarding (2 days)
  • ✓Dedicated Slack channel
  • ✓Quarterly architecture reviews
  • ✓Custom billing / PO support

All plans: cloud API calls (Anthropic, OpenAI, Gemini, etc.) pass through at exact provider cost — no Dream-Weaver markup. Cancel Pro anytime, no questions asked.

COMPARE

Full feature breakdown

Feature Open Core Pro Enterprise
CORE PROXY
OpenAI-compatible API ✓ ✓ ✓
Streaming (SSE) ✓ ✓ ✓
Function calling / tool use ✓ ✓ ✓
JSON mode + vision ✓ ✓ ✓
Embeddings API ✓ ✓ ✓
Local Ollama inference ✓ ✓ ✓
Cloud routing (Anthropic, OpenAI, Gemini…) ✓ ✓ ✓
INTELLIGENT ROUTING
Explicit model selection ✓ ✓ ✓
MoE auto-routing classifier – ✓ ✓
Policy-based routing rules – ✓ ✓
Routing audit log – ✓ ✓
Cost attribution by user / team – ✓ ✓
MEMORY & REASONING
Persistent semantic memory (pgvector) – ✓ ✓
Memory embedding model (mxbai-embed-large) – ✓ ✓
LSR neuro-symbolic reasoning pod – ✓ ✓
PLATFORM
MCP tool server – ✓ ✓
Multi-tenant API keys – ✓ ✓
Helm chart deployment ✓ ✓ ✓
SSO / SAML / LDAP – – ✓
Air-gapped / VPC deployment – – ✓
SLA uptime guarantee – – ✓
SUPPORT
Community (GitHub Issues) ✓ ✓ ✓
Email support – Same-day 4-hour SLA
Dedicated Slack channel – – ✓
Onboarding engineering – – ✓
FAQ

Questions answered.

Never. When you route a request to Anthropic, OpenAI, or Gemini, the call goes directly to their API using your API key. You pay them. We charge zero on top of that. Dream-Weaver's pricing covers the gateway infrastructure itself, not the tokens.
Dream-Weaver runs in your infrastructure — your Kubernetes cluster, your Docker host, your server. We never touch your traffic. There's no Dream-Weaver cloud middleman. Your data stays in your network. The Pro license key just enables feature flags in your local install.
Yes. Cancel anytime via Stripe — no questions, no lock-in period. You keep Pro features until the end of your billing period. After that, you revert to Open Core features, which are still fully functional for local inference and explicit routing.
No. Without a GPU, you can still route to cloud providers (Anthropic, OpenAI, Gemini, etc.) and get everything except local inference. A GPU (NVIDIA or AMD) is what gets you the 60–80% API cost reduction by running the cheap/fast tasks locally. Most modern RTX cards (12GB+ VRAM) work great.
Kubernetes 1.26+ is recommended. The Helm chart works with k3s, k3d, GKE, EKS, AKS, and vanilla kube. You can also run via Docker Compose without any Kubernetes at all. See the download page for system requirements.
After subscribing, you receive a license key (format: DW-PRO-XXXX-XXXX-XXXX) by email. Pass it to the Helm chart via --set license.key=DW-PRO-... or set DW_LICENSE_KEY in your Docker Compose environment. The proxy validates the key on startup and enables Pro feature flags. No data leaves your cluster for license validation.
A lightweight mixture-of-experts classifier that reads each incoming request and dispatches it to the optimal backend in <5ms. It categorizes requests by intent (code, reasoning, long-context, creative, fast-lookup) and applies affinity scores across your configured backends. Local GPU first for cheap tasks, frontier models for the hard ones. You configure the policy; the classifier makes the call.
GET STARTED

Start free today.

No credit card for Open Core. Pro unlocks the intelligence layer. Cancel anytime.

Download free Subscribe to Pro — $49/mo →

Enterprise pricing? Email us and we'll scope it out.