devwithjev
— reading now— views
Submit a build

Jev’s Four-Chef Kitchen Chaos

JollyRojak upgraded Jev’s Kitchen Chaos after V1, staggering requests so its four chefs can react independently. In a rerun, Jev scored 495, ahead of Claude Opus 4.8 (455) and GPT-5.6 Sol (450).

JollyRojak@Michael50663932𝕏
I upgraded Jev’s Kitchen Chaos after V1. Biggest change: the 4 chefs no longer make decisions at the exact same moment. Requests are now staggered, so each chef can react independently. Then I reran Jev vs Claude Opus 4.8 vs GPT-5.6 Sol. Results: 🥇 Jev: 495 🥈 Claude Opus 4.8: 455 🥉 GPT-5.6 Sol: 450 Watch the run 👇
Sep 19, 2026X postsView on X

Also filed under Games & real time

  • Benchmarks Jev on Pokémon Red

    The project benchmarks Jev, TypeSafe’s fast decision model, against Jev paired with a GPT-6 Sol planner for playing Pokémon Red.

  • Plain-English Cellular Ecosystem Simulator

    LifePot is a cellular automaton-inspired ecosystem that users describe in plain English. Jev turns those inputs into species, feeding relationships, reproduction strategies, and environmental conditions.

  • Mapping Children's Behavioral Intent in Games

    Mores Research used Jev to map children’s behavioral intent in games and examine how it correlates with real-world behavior. It reports mapping intents across 45+ real sessions involving 20 child users over a long horizon.

  • StarCraft Agent for Fast Tactical Decisions

    Jev_Star is a StarCraft agent experiment that uses Jev for fast tactical decisions.