Jev versus text-generating models
Jev does not generate free-form text. A side-by-side video compares parallel decision-making with token-by-token text generation.

A user ran Jev through FizzBuzz 10,000 times, finding a 77% failure rate, almost always on the same number, with Jev confidently treating 56 as divisible by 5 at 0.18-0.30 confidence.
I ran Jev through FizzBuzz 10,000 times. And it failed 77% of the time, and almost always on the exact same number. For whatever reason, Jev seems to really think that the number 56 is divisible by 5, and it does so with a confidence between 0.18-0.30. I have no idea what it