Video yükleniyor...
Video Yüklenemedi
Fable with low reasoning is roughly comparable to GPT-5.5 medium on the 20 puzzles they both solved. Fable scored higher in 14 of those, and I think Warming Tonic illustrates how different quality alchemical machines look. Solve rate isn't everything, quality is important!
41,084 görüntüleme • 2 ay önce •via X (Twitter)
0 Yorum
Yorum bulunmuyor
Orijinal gönderinin yorumları burada görünecek
