用 Jev 构建代码库分类器
开发者分享用 Jev 构建代码库分类器,认为这可能解决智能体生成过度工程化代码的问题,并询问下一步测试方向。
作者在Mac上运行8个本地模型,使用相同的JevBench问题测试。结果显示,4B模型在短文本上可匹配Jev,但复杂文本上所有本地模型均落后17分以上。
New article: how accurate are local models vs Jev, a cloud System One classifier? I ran 8 on a Mac, same JevBench questions. A 4B model matched Jev on short texts. On complex ones, all were 17+ points behind. https://rockyshikoku.medium.com/jev-style-text-classification-system-one-run-locally-how-accurate-can-it-get-57f30ed3594a… Video: minicpm5-2b classifying messages.