CONTROLAgents & automation
Jev vs. Local Models Snake Benchmark
This project benchmarks Jev against locally running Laya-MLX and Kev-4B models using a custom snake game, evaluating performance and output quality.
Jev vs locally running Laya-MLX and Kev-4B
I built my own snake benchmark. Jev @typesafeai is running via API.
The other models are small alternatives running on very little RAM on my Macbook!
Jev seems to deliver the best quality no doubt! After running it for a while 0 deaths.
Laya died pretty early on and then ran in a loop without catching any food.
Kev died twice in the same time but runs much faster at ~175ms latency while Jev needs ~395ms.
Laya is crazy fast at ~33ms but I guess you need to train it to increase output quality.
