Developer
Free
$0/mo
For development and initial workloads.
- Up to $15/mo monitored spend
- 3× sliding-window loop detection
- Hourly and daily budget caps
- OpenAI-compatible proxy endpoint
GATEWAY // PRICING & LIMITS
Start without a subscription and move to production controls when you need higher ceilings.
Loop savings calculator
$35/mo
Estimate uses the supplied $35 per prevented incident assumption multiplied by sessions per week. This is a planning estimate, not a guarantee or measured savings.
Developer
$0/mo
For development and initial workloads.
Production
$29/mo
For production agents that need higher budget ceilings.
Implementation details
The proxy is designed for under 35ms of average overhead. Actual latency depends on network distance, provider response time, and deployment conditions.
Prompt bodies are processed for loop detection in memory and are not written to disk. Request metadata and usage logs do not include prompt text.
Choose Zero-Trust Mode to send an upstream key per request; it is processed in memory and not stored. Encrypted Vault Mode uses AES-256 (Fernet) at rest. Shunt API keys are hashed for lookup.
A detected loop returns HTTP 200 with finish_reason set to stop. HTTP 400 indicates invalid request credentials, HTTP 429 indicates a budget ceiling, and HTTP 502 indicates an upstream provider error.
Any client or framework that can send OpenAI-compatible requests to a custom base URL can use the gateway, including CrewAI, LangChain, LangGraph, AutoGen, and direct REST clients.