Save money
Fast and coding lanes take cheap turns. Deep waits for hard work. Payload filtering drops stale history before it is billed.
OpenAI Chat Completions, OpenAI Responses, and Anthropic /v1/messages all classify each turn into a lane, trim waste context, and keep premium models in reserve. You pay OpenRouter for tokens. Strix sells the router and the memory layer.
Not a model catalog and not a token reseller. Strix is the decision layer between your coding clients and the providers you already pay.
Fast and coding lanes take cheap turns. Deep waits for hard work. Payload filtering drops stale history before it is billed.
Facts are extracted and injected silently across sessions and editors, scoped per project. Pin, recall, history, feedback, batch, and MCP on the same key. Toggle, export, or forget from the dashboard.
Builder or Team. Keys, seats after three, and spend visibility. Your OpenRouter key stays yours; Team is the SKU when more than one person routes.
Point Cursor, Claude Code, Codex, OpenCode, Cline, Roo Code, Continue, or Aider at one base URL. Strix classifies, filters, and egresses.
Classifies work into nine public intent models — coding, quick, deep, bug-hunt, visual, writing, research, uncensored — or leave the client on auto.
Removes stale context and oversized schemas while preserving the evidence your model actually needs.
Health-aware provider ordering, circuit state, streaming pass-through, and automatic transient failover.
OpenRouter BYOK is required. We do not mark up those tokens. Optional Venice only if you store a Venice key.
Agents resend history, tool schemas, and large tool results on every turn. A premium model then prices every retained token. Strix cuts both sides: send less, then route remaining work to the lowest-cost capable lane instead of always-Sonnet.
Drops stale history, duplicate context, oversized tool results, and idle schemas before they become billable input.
Quick and coding stay efficient. Advanced intelligence is reserved for difficult reasoning and escalation.
See tokens kept, tokens avoided, baseline cost, optimized cost, and savings by day, client, mode, and capability tier.
Change request volume and workload shape. The baseline uses the same Sonnet 4.6 list price used by the product evidence pipeline; optimized routes use prices already tracked in this repository.
At 5,000 requests per month, this scenario drops the modeled bill from $405 to $72.
27.2% measured Agent payload reduction in Strix telemetry. Uses 18,000 input and 1,800 output tokens per request. Estimate, not a guarantee; provider prices and workload mix change.
Pricing source: Anthropic Claude Platform pricing, accessed July 2026. Filter band source: repository evidence constants and measured Agent telemetry.
Strix does not summarize away the task. It preserves the current goal, relevant evidence, required tool protocol, and forced tools. It reduces categories of context that do not help this step.
“Refactor provider failover, preserve streaming, and add regression tests.”ROUTE → refactor-safe
Required tools and matched tool-call history remain attached.
Strix maintains a changing, health-checked range of capable model families. Your client keeps one configuration while routing adapts to the work, availability, and cost.
Fast utility work should not pay for Sol or Opus. Auto stays on cheap capable lanes. Frontier models stay in reserve unless you opt in.
You choose task-shaped lanes. Strix keeps provider identities, fallback order, and parameter compatibility behind the routing layer.
Create a workspace, copy your Strix key once, store OpenRouter (tokens stay on their bill), then point any supported client at the same base URL.
7-day trial · Builder $29/mo · Team $99/mo · OpenRouter BYOKBuilder for one person, Team when you need seats after three.
Tokens bill to OpenRouter. The full Strix key appears once — copy it then.
Cursor, Claude Code, Codex, OpenCode, Cline, Roo Code, Continue, or Aider. Same OpenAI-compatible contract.
7-day trial on both plans. No markup on BYOK tokens. Strix charges the subscription plus request and memory-derive overage.
7-day trial. You bring OpenRouter — they bill tokens; Strix does not mark up BYOK inference.
Three seats included, then $20 per extra seat. Same BYOK model: you pay OpenRouter for tokens.
Enter expected native OpenRouter usage. Builder is $29/mo; Team is $99/mo. Managed inference adds 15% only if Strix runs on the platform key (no stored OpenRouter credential).
BYOK totals are OpenRouter cost plus the Strix subscription. Managed inference adds 15% on native OpenRouter cost only when there is no stored key. Request and derive overage is metered separately.
Trial needs an OpenRouter key. Tokens stay on their bill. Strix is the router.