Back to explore

Jev의 의사결정형 에이전트와 평가 과제

Jev는 텍스트 생성이 아니라 도구 호출, 재시도, 라우팅, 에스컬레이션 같은 의사결정을 한다고 지적하며, 자율적 의사결정의 신뢰성 평가가 새로운 과제라고 주장한다.

@typesafeai Jev makes decisions instead of generating text: which tool to call, whether to retry, route or escalate Interesting direction for agents. Because once software makes millions of these decisions autonomously, “valid output” ≠ reliable decision. That’s an eval problem!

Jev의 의사결정형 에이전트와 평가 과제 1
· 1 likesOpen on X