Video wird geladen...
Video konnte nicht geladen werden
Fable with low reasoning is roughly comparable to GPT-5.5 medium on the 20 puzzles they both solved. Fable scored higher in 14 of those, and I think Warming Tonic illustrates how different quality alchemical machines look. Solve rate isn't everything, quality is important!
41,084 Aufrufe • vor 1 Monat •via X (Twitter)
0 Kommentare
Keine Kommentare verfügbar
Kommentare vom Original-Post werden hier angezeigt
