AnchorShell Relay
Control AI spend
before it gets out of hand.
AnchorShell Relay helps you spend less on AI by prioritizing lower-cost model paths, controlling requests, tokens, and spend, and showing exactly what every user and agent consumed.
Runaway-agent protection
Set the limit before the spend happens.
Control requests, tokens, or spend globally or for one user, agent, model, or provider. Apply the limit per second, minute, hour, day, or month.
- Reserve expensive models for the work that needs them.
- Give trusted workflows more room.
- Stop one agent without stopping everyone else.
Set a limit
Policy summary
Check the cost before the request runs.
Relay estimates incoming token usage before dispatch and checks it against configured token and spend limits.
Block new requests before their estimate exceeds the remaining allowance. Provider-reported usage can reconcile the estimate after completion when it is available.
Estimated token check
within limit
blocked before dispatch
Manage AI usage across the whole team.
AnchorShell Relay — Managed combines invitations, role presets, custom product permissions, request attribution, and per-user limits in one hosted operating model.
Managers can see organization usage while individual users can be limited to their own usage, logs, queue, and limits.
Team
AnchorShell Relay — ManagedTeam Members
3 activeUser controls
Dave · Developer- Requests
- 100 / day
- Tokens
- 100,000 / day
- Spend
- $5.00 / day
- Model-specific limits
- 3 configured
- Usage visibility
- Own usage
From queue to model to completed request.
Relay waits for configured capacity, skips paths that are cooling or outside their limits, and records the selected route with owner, token, cost, and timing data.
Manually rank local or free OpenAI-compatible paths first and keep an eligible cloud fallback ready.
Realtime request flow
One request moving through RelayRequest rq_214 waits in the queue, skips a cooling free endpoint, moves through the available local model, and completes with Maya as owner, 1,284 tokens, and four cents in cost.
Queue
Available models
Cooling path skipped
Completed requests
- Owner
- Maya
- Tokens
- 1,284
- Cost
- $0.04
Local model · 1.3s total
See usage before the invoice arrives.
See requests, tokens, and spend over time. Managed organization visibility can filter attributed usage by user alongside model, provider, and date range.
Start with the trend, narrow the scope, then inspect the exact routed request behind a spike, retry, or failure.
Usage over time
Requests, tokens, and spend across RelaySee exactly what your agent sent.
Operational metadata is recorded independently of payload capture. When deeper debugging is needed, an administrator with settings access can enable the global body-storage setting for new traffic.
Payload capture is optional and off by default. Relay does not require raw prompts and responses to retain request ownership, route, status, token, cost, and timing data.
Request history
Operational metadata for routed model calls| Request | Owner | Route | Result | Tokens | Timing | Cost |
|---|---|---|---|---|---|---|
| rq_7F2A | Coding agent | local/worker | Blocked | — | 42 ms | — |
| rq_7F2B | Maya | cloud/gpt-4.1 | Completed | 2,318 | 1.3 s | $0.04 |
| rq_7F2C | Research | free/general | Completed | 860 | 920 ms | $0.00 |
Store request and response bodies
Global Relay setting · off by defaultNew request and response bodies are not retained while capture is off.
Fewer failed requests. Less recovery code.
When enabled, Relay can move eligible requests after connection failures, timeouts, or upstream server failures. Throttled paths can cool while queued work waits for an available option.
Provider fallback
Configured model path- Primary modelcloud/primaryProvider statusUnhealthyskipped
- Regional backupcloud/secondaryProvider statusAvailableselected
- Local modellocal/workerProvider statusAvailableready
- Reserve modelcloud/reserveProvider statusAvailableready
Self-host Relay or let AnchorShell manage it.
Use AnchorShell Relay — Managed for organization-aware access and user attribution, or run AnchorShell Relay — Self-Hosted as a single-node service with SQLite-backed configuration and an embedded admin UI.