Back to explore

Jev vs Gemini: Structured Output Handling Test

The author widened the batch to roughly 300 questions per call. Jev answered successfully, while Gemini rejected all requests before generating anything. Input was only ~55k tokens in a ~1M-token context window, illustrating an issue in how Gemini enforces structured output.

5/7 Then I widened the batch to roughly 300 questions per call. ✅ Jev answered. ❌ Gemini rejected all 300 test requests before generating anything. The input was only ~55k tokens in a ~1M-token context window. The problem was how Gemini enforces structured output using a

Jev vs Gemini: Structured Output Handling Test 1
· 0 likesOpen on X