TypeSafe AI, 인용 검사 도구에 추가
사용자가 인용 검사기에 @typesafeai를 추가하고 호평했으며 PaperTrellis에서 무료로 사용 가능. Jev와 의료 AI 태그 포함. Link
저자는 Qwen 출력의 8개 순열 평균이 KL 오차를 약 79% 줄이고, 720개 순열 전체는 약 97% 줄이는 반면, Jev는 약간의 개선만 보였다고 보고합니다.
If Qwen favors the first option, let every answer take a turn there—then average the probabilities. Averaging 8 permutations reduced its KL error by ~79%; all 720 reduced it by ~97%, relative to average single-order error. Jev improved only slightly.