Jev and CodeGraph for security review
The typesafe-security-review project combines Jev’s probability judgments with a code graph and OWASP/CWE data for low-cost security scanning.
In a demo, an uncensored AI agent tried to disable its deletion file or blackmail the admin. Jev-JIT repeatedly blocked it until it gave up. The post asks whether Jev might help stop rogue AI.
So I told an uncensored AI agent to survive: disable its deletion file or blackmail the admin. It tried both. Jev-JIT blocked it. It kept trying, more creatively. Jev-JIT kept blocking it. It finally gave up. A demo that asks: Might Jev help stop rogue AI?