Back to explore

Independent Evaluation of Jev on AI Security Validation Sets

The author compares Jev with other models on their own AI security validation set, an independent hard benign benchmark, and a third-party hard benign benchmark, concluding Jev is far from bad but not production ready and not air gapped.

Because everybody here is comparing Jev with other models. I compared Jev on our #AISecurity Validation set, a independent hard benign benchmark, and a third party hard benign one. Result: #Jev is far from bad but not production ready. Also not air gapped.

Independent Evaluation of Jev on AI Security Validation Sets 1
· 0 likesOpen on X