Загрузка видео...
Не удалось загрузить видео
Fable with low reasoning is roughly comparable to GPT-5.5 medium on the 20 puzzles they both solved. Fable scored higher in 14 of those, and I think Warming Tonic illustrates how different quality alchemical machines look. Solve rate isn't everything, quality is important!
41,084 просмотров • 2 месяцев назад •via X (Twitter)
Комментарии: 0
Нет доступных комментариев
Здесь появятся комментарии из оригинального поста
