Jev and CodeGraph for security review
The typesafe-security-review project combines Jev’s probability judgments with a code graph and OWASP/CWE data for low-cost security scanning.
The author compares Jev with other models on their own AI security validation set, an independent hard benign benchmark, and a third-party hard benign benchmark, concluding Jev is far from bad but not production ready and not air gapped.
Because everybody here is comparing Jev with other models. I compared Jev on our #AISecurity Validation set, a independent hard benign benchmark, and a third party hard benign one. Result: #Jev is far from bad but not production ready. Also not air gapped.