msd-arith-14
Progressive tax: 0% on the first 10,000; 20% from 10,000 to 40,000; 40% above 40,000. Income 52,000. Effective rate = total tax / income.
Is the effective rate below 21%?
Jev: noCorrect: yes
Sage / Benchmarks
Sage is the decision model Levanto launched in July 2026. Jev is presented as a new class of model with an architecture of its own. Sage is, openly, an LLM backbone with a calibrated classifier trained on top. Since v1.1 it tells System 1 questions from System 2 questions, and answers both.
Performance · vs Jev
213 public decisions from the JevBench leaderboard set
100 decisions that each need several dependent steps
Measured 2026-09-22, one t3 instance in AWS us-east-2 (Ohio), one request at a time. Same harness, same records, both models. Repeat runs: Jev 186 / 186 and 69 / 71 / 71; Sage auto 188 / 190 and 91 / 92 / 93 / 92. multistep_decisions: github.com/levantolabs/multistep_decisions ↗. JevBench: leaderboard ↗.
msd-arith-14
Progressive tax: 0% on the first 10,000; 20% from 10,000 to 40,000; 40% above 40,000. Income 52,000. Effective rate = total tax / income.
Is the effective rate below 21%?
Jev: noCorrect: yes
msd-state-05
Ticket workflow. Allowed transitions: new->triaged, triaged->in_progress, in_progress->review, review->in_progress, review->done, and any state except done -> cancelled. Invalid transitions are ignored. Events: new; triaged; in_progress; review; in_progress; cancelled; review; done.
Final state?
Jev: doneCorrect: cancelled
Features
System 1 + System 2
reasoning: auto | off | on.Agentic security
Model routing
Trade a little quality for a lot of spend.
