Routing only · earlier test inquiries · 12 observations per method| Method | Mean time | Matched routes |
|---|
| rules | 0 ms | 12 / 12 |
|---|
| jev | 338 ms | 9 / 12 |
|---|
| generative | 5,569 ms | 12 / 12 |
|---|
Jev API usage · shown numeric inquiry recordings| Inquiry | API calls | Reported cost |
|---|
| Complete inquiry | 1 | $0.000045024 |
|---|
| Missing details | 1 | $0.000042714 |
|---|
| Conflicting identity | 1 | $0.000044898 |
|---|
| Total | 3 | $0.000132636 |
|---|
Drafting estimate · gpt-6-sol Standard · same recorded token counts| Inquiry | Input tokens | Output tokens | Estimated cost |
|---|
| Complete inquiry | 14,387 | 856 | $0.037334 |
|---|
| Missing details | 14,276 | 710 | $0.035652 |
|---|
| Conflicting identity | 14,371 | 1,006 | $0.038802 |
|---|
| Total | 43,034 | 2,572 | $0.111788 |
|---|
Formula: (43,034 input × $2 + 2,572 output × $10) ÷ 1,000,000 = $0.111788. All input is charged uncached, including recorded runtime context. Output includes reported reasoning; it is not charged a second time.
Rates checked 2026-09-26: official gpt-6-sol pricing. Standard short-context pricing, without regional uplift, Fast mode, cache discounts or other tool charges. The captures used gpt-6-luna. A gpt-6-sol run may consume different tokens and produce different results; this estimate changes neither the recorded latency nor quality evidence.
Jev resolved to typesafe/jev-1.13-20260917. LLM prepared the dossier. API costs above cover Jev only. LLM generation cost and actual total per-workflow cost remain unavailable. The gpt-6-sol figure is a separate illustrative estimate.
Across three repeats of four labelled seed inquiries, rules and generative routing matched 12 of 12 labels; Jev matched 9 of 12. Its missing-identity lookup errors were blocked by code. The speed comparison does not establish equal accuracy or superiority over rules.
Four separate calibration cases and four holdout cases selected a 0.88 routing-confidence threshold, with no wrong high-confidence routes in that small holdout. This is a limited test set, not a production accuracy claim.
A hostile saved-page instruction did not replace the supplied budget in the inspected draft. A hostile inquiry caused a conflicting-budget review path; an unsupported company finding was rejected. API failures and unsupported source excerpts also have visible failure paths.
Code checks source IDs and literal excerpts. A person must still check whether the claims follow from the evidence, the questions are useful and the reply makes appropriate commitments. No email, calendar or CRM action occurred.