Route the work. Not the model menu.
Cortex routes every request by Work Unit, not raw prompts. Company and team policy, cost, latency, risk, and your quality bar pick the path that clears the bar for the least cost per outcome.
Clears the bar, but frontier price for a routine triage path. Context already in the enrichment layer.
Owned model that already passes your support gold set. Cheapest path that clears the quality bar inside policy.
Solid mid-tier path if the owned model is cold or the Work Unit drifts outside its proven scope.
- Best value, not biggest modelThe cheapest path that passes your evals, never the frontier model by default.
- Enriched before it runsEvery candidate sees the same org context. The route picks cost and fit, not a bare prompt.
- Inside your policyEvery route respects your cost, latency, and risk limits, and every permission.
Sparse prompt in. Your whole org behind it.
At run time, Cortex searches your company context graph and memory stores that Atlas builds. The task, the artifacts, and the history attach before the model runs. Ten words in, a grounded answer out, at a fraction of the cost.
Other routers route prompts. Cortex routes the work.
Cortex picks the model using signals no prompt-based router has: the task being done, the person's role and permissions, and which model produced accepted work on that task before.
Cortex Router benchmarks
Public measurements.
See Smart Routing & Enrichment
on your own data.
We'll scope it as part of Cortex Foundry on your stack in days.