devwithjev
— reading now— views
Submit a build

Plays Tetris by Selecting Moves

Jev played Tetris in real time using the same 200-piece sequence and legal move options as Claude Haiku 4.5 and Gemini 3.5 Flash-Lite. It scored 9,200 points, cleared 75 lines, averaged 300 milliseconds per move, and made zero errors.

View on X time平均每步 300 毫秒
Sliven 褚崇名@SlivenRed𝕏
同一組 200 個方塊、相同的合法移動選項,讓 Jev、Claude Haiku 4.5 和 Gemini 3.5 Flash-Lite 即時玩俄羅斯方塊。 Jev:9,200 分、消除 75 行,平均每步 300 毫秒,0 錯誤。 Claude:8,900 分、消除 72 行,平均每步 1.52 秒,0 錯誤,花費 US$0.48。 Gemini:9,000 分、消除 75 行,平均每步 1.13 秒,0 錯誤,花費 US$0.02。 這次 Jev 拿下最高分,每步決策速度約是另外兩個模型的 4~5 倍。 這不代表 Jev 更聰明。當任務只是從合法選項裡挑出下一步,不一定需要模型生成一段回答。每次決策少做一些工作,就可能更快、更便宜。 https://t.co/oPcnofcrug
Sep 18, 2026X postsView on X
Jev had the highest score in the comparison, with decisions roughly 4–5 times faster than the other two models. The author cautions that this does not mean Jev is smarter; choosing among legal moves may require less work than generating a response.

Also filed under Games & real time

  • Steer a 3D Box with Goals

    The build steers a 3D box with plain-language goals. An LLM plans each instruction, while TypeSafe Jev makes move, turn, and jump decisions in real time.

  • Benchmarks Jev on Pokémon Red

    The project benchmarks Jev, TypeSafe’s fast decision model, against Jev paired with a GPT-6 Sol planner for playing Pokémon Red.

  • Plain-English Cellular Ecosystem Simulator

    LifePot is a cellular automaton-inspired ecosystem that users describe in plain English. Jev turns those inputs into species, feeding relationships, reproduction strategies, and environmental conditions.

  • Mapping Children's Behavioral Intent in Games

    Mores Research used Jev to map children’s behavioral intent in games and examine how it correlates with real-world behavior. It reports mapping intents across 45+ real sessions involving 20 child users over a long horizon.