devwithjev
โ€” reading nowโ€” views
Submit a build

Routes Review Tasks with Jev

Jev was tested on 32 real historical tasks. It identified 103/126 required reviewers, had 0/41 critical misses, and caught 7/7 escalation cases.

Pranay Suyash@pranaysuyash๐•
Yesterday I posted the Jev routing experiment before I had API access. Got access later and ran it on 32 real historical tasks. Three-way comparison: Current council: 125/126 required reviewers, 0/41 critical misses, 1/7 escalation cases caught. Jev: 103/126 required reviewers, 0/41 critical misses, 7/7 escalation cases caught. Typed GPT-4.1-mini control, with the same state and the same 21 questi
Sep 21, 2026X postsView on X
The author ran the experiment after getting API access. For comparison, the current council identified 125/126 required reviewers, had 0/41 critical misses, and caught 1/7 escalation cases.

Also filed under Triage & routing