Back to explore

เอเจนต์ตัดสินใจของ Jev และความท้าทายด้านการประเมิน

โพสต์ระบุว่า Jev ตัดสินใจแทนการสร้างข้อความ เช่น เรียกใช้เครื่องมือ ลองใหม่ กำหนดเส้นทาง หรือยกระดับ และโต้แย้งว่าการประเมินการตัดสินใจอัตโนมัติที่เชื่อถือได้เป็นความท้าทายใหม่

@typesafeai Jev makes decisions instead of generating text: which tool to call, whether to retry, route or escalate Interesting direction for agents. Because once software makes millions of these decisions autonomously, “valid output” ≠ reliable decision. That’s an eval problem!

เอเจนต์ตัดสินใจของ Jev และความท้าทายด้านการประเมิน 1
· 1 likesOpen on X