Developer
Free
$0/mo
For development and initial workloads.
- Up to $15/mo monitored spend
- 3× sliding-window loop detection
- Hourly and daily budget caps
- OpenAI-compatible proxy endpoint
GATEWAY // PRICING & LIMITS
Start without a subscription and move to production controls when you need higher ceilings.
Loop savings calculator
$35/mo
Estimate uses the supplied $35 per prevented incident assumption multiplied by sessions per week. This is a planning estimate, not a guarantee or measured savings.
Developer
$0/mo
For development and initial workloads.
Production
$29/mo
For production agents that need higher budget ceilings.
Implementation details
The proxy is designed for under 35ms of average overhead. Actual latency depends on network distance, provider response time, and deployment conditions.
Prompt bodies are processed for loop detection in memory and are not written to disk. Request metadata and usage logs do not include prompt text.
Upstream credentials are encrypted with AES-256 (Fernet) before database insertion. Protected key material is hashed for lookup; raw upstream keys are not logged.
HTTP 400 indicates an infinite loop was stopped, HTTP 429 indicates a budget ceiling was reached, and HTTP 502 indicates an upstream provider error.
Any client or framework that can send OpenAI-compatible requests to a custom base URL can use the gateway, including CrewAI, LangChain, LangGraph, AutoGen, and direct REST clients.